---
title: HybridRAG: Integrating Knowledge Graphs and Vector Retrieval Augmented Generation for Efficient Information Extraction
url: https://www.ml-quant.com/papers/arxiv/2408.04948/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2408.04948
source_url: https://arxiv.org/abs/2408.04948
featured: 2024-08-15
citations: 231
topic: LLMs & Text
---


# HybridRAG: Integrating Knowledge Graphs and Vector Retrieval Augmented Generation for Efficient Information Extraction

Q&A Systems for Financial Data: HybridRAG, a new method combining Knowledge Graphs and VectorRAG techniques, improves question-answer systems for extracting information from financial documents, offering better accuracy and answer generation.

- Source: https://arxiv.org/abs/2408.04948
- Identifier: arXiv:2408.04948
- Released: 2024-08-09
- First featured: Quant Letter No. 61 (2024-08-15): https://www.ml-quant.com/issues/2024-08-15/
- Citations (Semantic Scholar): 231
- Published in: Proceedings of the 5th ACM International Conference on AI in Finance
- Topic: LLMs & Text

## Related

- [RAFT: Adapting Language Model to Domain Specific RAG](https://www.ml-quant.com/papers/arxiv/2403.10131/): Retrieval Augmented FineTuning (RAFT) is a new training method that enhances large language models' ability to answer domain-specific questions by training them to ignore irrelevant documents and cite relevant ones.
- [ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities](https://www.ml-quant.com/papers/arxiv/2407.14482/): Bridging the Gap: ChatQA 2 is a model that improves long-context understanding and retrieval-augmented generation, matching the accuracy of top proprietary models.
- [RAGGED: Towards Informed Design of Scalable and Stable RAG Systems](https://www.ml-quant.com/papers/arxiv/2403.09040/): RAGGED, a new framework, optimizes language models for document-based question answering by analyzing Retrieval-augmented generation configurations.
- [KnowPO: Knowledge-aware Preference Optimization for Controllable Knowledge Selection in Retrieval-Augmented Language Models](https://www.ml-quant.com/papers/arxiv/2408.03297/): The study introduces a Knowledge-aware Preference Optimization method to improve large language models' knowledge selection, showing enhanced performance in managing knowledge conflicts and robust generalization across different datasets.
- [FACTS About Building Retrieval Augmented Generation-based Chatbots](https://www.ml-quant.com/papers/arxiv/2407.07858/): The article introduces the FACTS framework for developing Retrieval Augmented Generation (RAG)-based chatbots, and presents empirical results on the balance between accuracy and latency in large and small LLMs.
- [From RAGs to rich parameters: Probing how language models utilize external knowledge over parametric information for factual queries](https://www.ml-quant.com/papers/arxiv/2406.12824/): Retrieval Augmented Generation (RAG) enhances language models' reasoning abilities using external context, but models tend to rely heavily on this context and less on their parametric memory.
