Mix-of-Granularity: Optimize the Chunking Granularity for   Retrieval-Augmented Generation

Zijie Zhong; Hanwen Liu; Xiaoya Cui; Xiaofan Zhang; Zengchang Qin

arXiv:2406.00456·cs.LG·January 28, 2025·3 cites

Mix-of-Granularity: Optimize the Chunking Granularity for Retrieval-Augmented Generation

Zijie Zhong, Hanwen Liu, Xiaoya Cui, Xiaofan Zhang, Zengchang Qin

PDF

Open Access 1 Repo

TL;DR

This paper introduces Mix-of-Granularity (MoG), a dynamic method for optimizing knowledge source chunking in Retrieval-Augmented Generation systems, improving information retrieval and downstream task performance.

Contribution

The paper proposes MoG and MoG-Graph, novel methods that adaptively determine the best data granularity for knowledge retrieval, with a new training loss and graph-based extension.

Findings

01

MoG accurately predicts optimal granularity levels.

02

MoGG enhances retrieval of distant snippets.

03

Both methods improve RAG system performance.

Abstract

Integrating information from various reference databases is a major challenge for Retrieval-Augmented Generation (RAG) systems because each knowledge source adopts a unique data structure and follows different conventions. Retrieving from multiple knowledge sources with one fixed strategy usually leads to under-exploitation of information. To mitigate this drawback, inspired by Mix-of-Expert, we introduce Mix-of-Granularity (MoG), a method that dynamically determines the optimal granularity of a knowledge source based on input queries using a router. The router is efficiently trained with a newly proposed loss function employing soft labels. We further extend MoG to MoG-Graph (MoGG), where reference documents are pre-processed as graphs, enabling the retrieval of distantly situated snippets. Experiments demonstrate that MoG and MoGG effectively predict optimal granularity levels,…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

ZGChung/Mix-of-Granularity
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsNatural Language Processing Techniques

MethodsRefunds@Expedia|||How do I get a full refund from Expedia? · Attention Is All You Need · Layer Normalization · Dense Connections · Adam · Softmax · Linear Warmup With Linear Decay · Residual Connection · Dropout · Byte Pair Encoding