Search papers, labs, and topics across Lattice.
This paper introduces SAG (SQL-Retrieval Augmented Generation), a novel architecture that enhances Retrieval-Augmented Generation by utilizing dynamic hyperedges to link semantically complete events at query time, rather than relying on a pre-built global graph. This approach addresses the limitations of existing methods in handling structured constraints and multi-hop reasoning while minimizing maintenance overhead and allowing for incremental updates. SAG outperforms existing systems on multiple multi-hop benchmarks, achieving an impressive 80.0% Recall@5 on MuSiQue, and demonstrates practical scalability with rapid online retrieval capabilities.
SAG achieves 80.0% Recall@5 on the challenging MuSiQue benchmark, showcasing a breakthrough in multi-hop reasoning without the burden of global graph maintenance.
Retrieval-Augmented Generation (RAG) offers an effective approach for large language models to access external knowledge. However, existing methods rely on dense similarity retrieval and face inherent limitations in handling structured constraints and multi-hop reasoning. Incorporating knowledge graphs partially alleviates these issues, but at the cost of semantic fragmentation, high maintenance overhead, and difficult incremental updates. This paper introduces SAG (SQLRetrieval Augmented Generation), a structured architecture for retrieval and agent systems. Instead of pre-building a global static graph, SAG converts each chunk into one semantically complete event and a set of indexing entities, then uses SQL join queries to dynamically link events that share entities into local hyperedges,constructing, at query time, a dynamically instantiated local index structure. This design avoids the need for global graph rebuilding and ongoing maintenance; the system naturally supports incremental writes, concurrent processing, and continuous scaling through its reliance on standard database infrastructure. Across HotpotQA, 2WikiMultiHop, and MuSiQue, three standard multi-hop benchmarks,SAG achieves the best results on 8 out of 9 Recall@K metrics, reaching 80.0% Recall@5 on MuSiQue, the benchmark with the highest multi-hop reasoning demands.SAG has also been deployed at a production scale of hundreds of millions of data items, with online retrieval latency kept within seconds. Project site and code are available at https://github.com/Zleap-AI/SAG-Benchmark.