LLM Rag
Curates a categorized list of papers, models, tools, and projects for retrieval-augmented generation (RAG) research and development.
What this file does
Curates a categorized list of papers, models, tools, and projects for retrieval-augmented generation (RAG) research and development.
When to use it
- Surveying recent RAG papers and benchmarks
- Finding open-source RAG projects and libraries
- Comparing embedding models and vector databases
- Exploring multimodal and agentic RAG approaches
LLM Rag
Survey
-
Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs,
arXiv, 2507.09477, arxiv, pdf, cication: -1Yangning Li, Weizhi Zhang, Yuyao Yang, ..., Yangqiu Song, Philip S. Yu ยท (Awesome-RAG-Reasoning - DavidZWZ)
-
Ask in Any Modality: A Comprehensive Survey on Multimodal Retrieval-Augmented Generation,
arXiv, 2502.08826, arxiv, pdf, cication: -1Mohammad Mahdi Abootorabi, Amirhosein Zobeiri, Mahdi Dehghani, ..., Mahdieh Soleymani Baghshah, Ehsaneddin Asgari ยท (Multimodal-RAG-Survey - llm-lab-org)
-
Towards Trustworthy Retrieval Augmented Generation for Large Language Models: A Survey,
arXiv, 2502.06872, arxiv, pdf, cication: -1Bo Ni, Zheyuan Liu, Leyao Wang, ..., Meng Jiang, Tyler Derr
RAG
-
MetaEmbed: Scaling Multimodal Retrieval at Test-Time with Flexible Late Interaction,
arXiv, 2509.18095, arxiv, pdf, cication: -1Zilin Xiao, Qi Ma, Mengting Gu, ..., Vicente Ordonez, Vijai Mohan
-
๐ Vision-Guided Chunking Is All You Need: Enhancing RAG with Multimodal Document Understanding,
arXiv, 2506.16035, arxiv, pdf, cication: -1Vishesh Tripathi, Tanmay Odapally, Indraneel Das, ..., Uday Allu, Biddwan Ahmed
-
DoTA-RAG: Dynamic of Thought Aggregation RAG,
arXiv, 2506.12571, arxiv, pdf, cication: -1Saksorn Ruangtanusak, Natthapath Rungseesiripak, Peerawat Rojratchadakorn, ..., Monthol Charattrakool, Natapong Nitarach
-
UniversalRAG: Retrieval-Augmented Generation over Corpora of Diverse Modalities and Granularities,
arXiv, 2504.20734, arxiv, pdf, cication: -1Woongyeong Yeo, Kangsan Kim, Soyeong Jeong, ..., Jinheon Baek, Sung Ju Hwang ยท (universalrag.github)
-
ReasonIR: Training Retrievers for Reasoning Tasks,
arXiv, 2504.20595, arxiv, pdf, cication: -1Rulin Shao, Rui Qiao, Varsha Kishore, ..., Pang Wei Koh, Luke Zettlemoyer
-
๐ NodeRAG: Structuring Graph-based RAG with Heterogeneous Nodes,
arXiv, 2504.11544, arxiv, pdf, cication: -1Tianyang Xu, Haojie Zheng, Chengze Li, ..., Ruoxi Chen, Lichao Sun ยท (NodeRAG. - Terry-Xu-666)
-
RouterRetriever: Routing over a Mixture of Expert Embedding Models,
arXiv, 2409.02685, arxiv, pdf, cication: -1Hyunji Lee, Luca Soldaini, Arman Cohan, ..., Minjoon Seo, Kyle Lo ยท (๐)
-
Enhancing Financial Time-Series Forecasting with Retrieval-Augmented Large Language Models,
arXiv, 2502.05878, arxiv, pdf, cication: -1Mengxi Xiao, Zihao Jiang, Lingfei Qian, ..., Sophia Ananiadou, Qianqian Xie ยท (huggingface)
-
SafeRAG: Benchmarking Security in Retrieval-Augmented Generation of Large Language Model,
arXiv, 2501.18636, arxiv, pdf, cication: -1Xun Liang, Simin Niu, Zhiyu Li, ..., Mengwei Wang, Jiawei Yang ยท (SafeRAG - IAAR-Shanghai)
-
SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models,
arXiv, 2502.09604, arxiv, pdf, cication: -1Yung-Sung Chuang, Benjamin Cohen-Wang, Shannon Zejiang Shen, ..., Shang-Wen Li, Wen-tau Yih
-
Chain-of-Retrieval Augmented Generation,
arXiv, 2501.14342, arxiv, pdf, cication: -1Liang Wang, Haonan Chen, Nan Yang, ..., Zhicheng Dou, Furu Wei
-
MiniRAG: Towards Extremely Simple Retrieval-Augmented Generation,
arXiv, 2501.06713, arxiv, pdf, cication: -1Tianyu Fan, Jingyuan Wang, Xubin Ren, ..., Chao Huang ยท (MiniRAG - HKUDS)
-
๐ OmniThink: Expanding Knowledge Boundaries in Machine Writing through Thinking,
arXiv, 2501.09751, arxiv, pdf, cication: -1Zekun Xi, Wenbiao Yin, Jizhan Fang, ..., Fei Huang, Huajun Chen ยท (zjunlp.github)
-
Knowledge Models Combine Retrieval with Generation: An Introduction to RAG
-
Personalized Graph-Based Retrieval for Large Language Models,
arXiv, 2501.02157, arxiv, pdf, cication: -1Steven Au, Cameron J. Dimacali, Ojasmitha Pedirappagari, ..., Ryan A. Rossi, Nesreen K. Ahmed
-
GeAR: Generation Augmented Retrieval,
arXiv, 2501.02772, arxiv, pdf, cication: -1Haoyu Liu, Shaohan Huang, Jianfeng Liu, ..., Furu Wei, Qi Zhang
-
Don't Do RAG: When Cache-Augmented Generation is All You Need for Knowledge Tasks,
arXiv, 2412.15605, arxiv, pdf, cication: -1Brian J Chan, Chao-Ting Chen, Jui-Hung Cheng, ..., Hen-Hsen Huang ยท (cag - hhhuang)
-
Long Context vs. RAG for LLMs: An Evaluation and Revisits,
arXiv, 2501.01880, arxiv, pdf, cication: -1Xinze Li, Yixin Cao, Yubo Ma, ..., Aixin Sun
-
SKETCH: Structured Knowledge Enhanced Text Comprehension for Holistic Retrieval,
arXiv, 2412.15443, arxiv, pdf, cication: -1Aakash Mahalingam, Vinesh Kumar Gande, Aman Chadha, ..., Vinija Jain, Divya Chaudhary
-
GemmaEmbed is a dense-vector embedding model, trained especially for retrieval. ๐ค
-
๐ Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference,
arXiv, 2412.13663, arxiv, pdf, cication: -1Benjamin Warner, Antoine Chaffin, Benjamin Claviรฉ, ..., Jeremy Howard, Iacopo Poli ยท (huggingface) ยท (๐)
-
Auto-RAG: Autonomous Retrieval-Augmented Generation for Large Language Models,
arXiv, 2411.19443, arxiv, pdf, cication: -1Tian Yu, Shaolei Zhang, Yang Feng ยท (Auto-RAG - ictnlp)
-
ยท (arxiv)
-
Jina CLIP v2: Multilingual Multimodal Embeddings for Texts and Images ๐ค
ยท (huggingface)
-
Long Term Memory: The Foundation of AI Self-Evolution,
arXiv, 2410.15665, arxiv, pdf, cication: -1Xun Jiang, Feng Li, Han Zhao, ..., Mengdi Wang, Tianqiao Chen ยท (๐)
-
๐ HtmlRAG: HTML is Better Than Plain Text for Modeling Retrieved Knowledge in RAG Systems,
arXiv, 2411.02959, arxiv, pdf, cication: -1Jiejun Tan, Zhicheng Dou, Wen Wang, ..., Weipeng Chen, Ji-Rong Wen
-
M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding,
arXiv, 2411.04952, arxiv, pdf, cication: -1Jaemin Cho, Debanjan Mahata, Ozan Irsoy, ..., Yujie He, Mohit Bansal ยท (m3docrag.github)
-
In Defense of RAG in the Era of Long-Context Language Models,
arXiv, 2409.01666, arxiv, pdf, cication: 3Tan Yu, Anbang Xu, Rama Akkiraju ยท (zyphra)
-
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation,
arXiv, 2410.09584, arxiv, pdf, cication: -1Guanting Dong, Xiaoshuai Song, Yutao Zhu, ..., Zhicheng Dou, Ji-Rong Wen ยท (FollowRAG.github) ยท (arxiv) ยท (FollowRAG - dongguanting)
ยท (huggingface)
-
Meta-Chunking: Learning Efficient Text Segmentation via Logical Perception,
arXiv, 2410.12788, arxiv, pdf, cication: -1Jihao Zhao, Zhiyuan Ji, Pengnian Qi, ..., Feiyu Xiong, Zhiyu Li ยท (Meta-Chunking - IAAR-Shanghai)
ยท (arxiv)
-
Your Mixture-of-Experts LLM Is Secretly an Embedding Model For Free,
arXiv, 2410.10814, arxiv, pdf, cication: -1Ziyue Li, Tianyi Zhou ยท (MoE-Embedding - tianyi-lab)
Multi Modal
-
๐ VideoRAG: Retrieval-Augmented Generation over Video Corpus,
arXiv, 2501.05874, arxiv, pdf, cication: -1Soyeong Jeong, Kangsan Kim, Jinheon Baek, ..., Sung Ju Hwang ยท (huggingface) ยท (๐)
-
Visual Document Retrieval Goes Multilingual ๐ค
ยท (๐)
-
MM-Embed, an extension of NV-Embed-v1 with multimodal retrieval capability. ๐ค
-
Beyond Text: Optimizing RAG with Multimodal Inputs for Industrial Applications,
arXiv, 2410.21943, arxiv, pdf, cication: -1Monica Riedler, Stefan Langer ยท (x)
-
VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents,
arXiv, 2410.10594, arxiv, pdf, cication: -1Shi Yu, Chaoyue Tang, Bokai Xu, ..., Zhiyuan Liu, Maosong Sun
-
MMed-RAG: Versatile Multimodal RAG System for Medical Vision Language Models,
arXiv, 2410.13085, arxiv, pdf, cication: -1Peng Xia, Kangyu Zhu, Haoran Li, ..., James Zou, Huaxiu Yao
Embedding
-
Breaking the Modality Barrier: Universal Embedding Learning with Multimodal LLMs,
arXiv, 2504.17432, arxiv, pdf, cication: -1Tiancheng Gu, Kaicheng Yang, Ziyong Feng, ..., Weidong Cai, Jiankang Deng
-
nomic-embed-text-v2-moe: Multilingual Mixture of Experts Text Embeddings ๐ค
-
Train 400x faster Static Embedding Models with Sentence Transformers ๐ค
ยท (๐)
Evaluation
-
WebWalker: Benchmarking LLMs in Web Traversal,
arXiv, 2501.07572, arxiv, pdf, cication: -1Jialong Wu, Wenbiao Yin, Yong Jiang, ..., Pengjun Xie, Fei Huang ยท (WebWalker - Alibaba-NLP)
-
๐ MMDocIR: Benchmarking Multi-Modal Retrieval for Long Documents,
arXiv, 2501.08828, arxiv, pdf, cication: -1Kuicai Dong, Yujing Chang, Xin Deik Goh, ..., Ruiming Tang, Yong Liu
-
OmniEval: An Omnidirectional and Automatic RAG Evaluation Benchmark in Financial Domain,
arXiv, 2412.13018, arxiv, pdf, cication: -1Shuting Wang, Jiejun Tan, Zhicheng Dou, ..., Ji-Rong Wen ยท (OmniEval - RUC-NLPIR)
-
Long Context RAG Performance of Large Language Models,
arXiv, 2411.03538, arxiv, pdf, cication: -1Quinn Leng, Jacob Portes, Sam Havens, ..., Matei Zaharia, Michael Carbin
-
CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation,
arXiv, 2410.23090, arxiv, pdf, cication: -1Yiruo Cheng, Kelong Mao, Ziliang Zhao, ..., Ji-Rong Wen, Zhicheng Dou ยท (CORAL - Ariya12138)
Database
Projects
-
memvid - Olow304
-
RAG-Anything - HKUDS
All-in-One RAG System
-
all-rag-techniques - FareedKhan-dev
A Simpler, Hands-On Approach โจ
-
VRAG - Alibaba-NLP
Moving Towards Next-Generation RAG via Multi-Modal Agent RL
-
Finetune-Bench-RAG - Pints-AI
Fine-tuning Models to Tackle Retrieval-Augmented Generation (RAG) Hallucination
-
chatwiki - zhimaAi
-
graphiti - getzep
-
RAG_Techniques - NirDiamant
Elevating Your Retrieval-Augmented Generation Systems ๐
-
LightRAG - HKUDS
Simple and Fast Retrieval-Augmented Generation
-
pathway - pathwaycom
-
onyx - onyx-dot-app
-
fast-graphrag - circlemind-ai
-
txtai - neuml
-
Perplexica - ItzCrazyKns
-
dsRAG - D-Star-AI
-
๐ RAGViz - cxcscmu
ยท (youtube)
-
pgai - timescale
-
Contextual RAG from Anthropic ๐
ยท (together-cookbook - togethercomputer)
-
AutoRAG - Marker-Inc-Korea
-
KAG - OpenSPG
Knowledge Augmented Generation
Products
Misc
-
GraphRAG-esque metadata tagging + retrieval ๐
ยท (llama_parse - run-llama)
-
ยท (๐)
-
Expert Support case study: Bolstering a RAG app with LLM-as-a-Judge ๐ค
Vector Database
What's inside
10 sections: Survey, RAG, Multi Modal, Embedding, Evaluation, Database, Projects, Products, Misc, Vector Database
Change this for your project
- Replace
metame-ai/awesome-llm-plazawith your own repository name if forking - Replace
DavidZWZ/Awesome-RAG-Reasoningwith your own survey repo link if maintaining a fork - Replace
HKUDS/MiniRAGwith your own project reference if customizing the list
Where it goes
Reference documentation for a retrieval pipeline. Keep with the ingestion or retrieval code it describes.
Worth borrowing
- Grouping resources by sub-topic (e.g., Multi Modal, Evaluation) for quick scanning
- Using star badges and external links (arXiv, Hugging Face, GitHub) to surface quality and provenance
Related Documents
SUMMARY
Proposes three on-prem AI architectures, modular, hybrid, and fully local RAG, with hardware specs and vendor lists.
Retrieval & Prompts
Explains how CharMemory's extraction prompt and Vector Storage settings determine memory retrieval quality in SillyTavern.
App Review Support Guide โ Switch2Go
Explains an AAC app's accessibility permissions, hardware needs, and reviewer walkthrough to pass App Store review.
RFC-BLite: High-Performance Embedded Document Database for .NET
Specifies an embedded document database for.NET with zero-allocation I/O, C-BSON format, and ACID transactions.