Back to .md Directory

LLM Rag

Curates a categorized list of papers, models, tools, and projects for retrieval-augmented generation (RAG) research and development.

May 2, 2026
0 downloads
0 views
ai agent llm rag eval
View source

What this file does

Curates a categorized list of papers, models, tools, and projects for retrieval-augmented generation (RAG) research and development.

When to use it

  • Surveying recent RAG papers and benchmarks
  • Finding open-source RAG projects and libraries
  • Comparing embedding models and vector databases
  • Exploring multimodal and agentic RAG approaches

LLM Rag

Survey

  • Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs, arXiv, 2507.09477, arxiv, pdf, cication: -1

    Yangning Li, Weizhi Zhang, Yuyao Yang, ..., Yangqiu Song, Philip S. Yu ยท (Awesome-RAG-Reasoning - DavidZWZ) Star

  • Ask in Any Modality: A Comprehensive Survey on Multimodal Retrieval-Augmented Generation, arXiv, 2502.08826, arxiv, pdf, cication: -1

    Mohammad Mahdi Abootorabi, Amirhosein Zobeiri, Mahdi Dehghani, ..., Mahdieh Soleymani Baghshah, Ehsaneddin Asgari ยท (Multimodal-RAG-Survey - llm-lab-org) Star

  • Towards Trustworthy Retrieval Augmented Generation for Large Language Models: A Survey, arXiv, 2502.06872, arxiv, pdf, cication: -1

    Bo Ni, Zheyuan Liu, Leyao Wang, ..., Meng Jiang, Tyler Derr

RAG

  • MetaEmbed: Scaling Multimodal Retrieval at Test-Time with Flexible Late Interaction, arXiv, 2509.18095, arxiv, pdf, cication: -1

    Zilin Xiao, Qi Ma, Mengting Gu, ..., Vicente Ordonez, Vijai Mohan

  • ๐ŸŒŸ Vision-Guided Chunking Is All You Need: Enhancing RAG with Multimodal Document Understanding, arXiv, 2506.16035, arxiv, pdf, cication: -1

    Vishesh Tripathi, Tanmay Odapally, Indraneel Das, ..., Uday Allu, Biddwan Ahmed

  • Multi-Agent ๅไฝœๅ…ด่ตท๏ผŒRAG ๆณจๅฎšๅชๆ˜ฏ่ฟ‡ๆธกๆ–นๆกˆ๏ผŸ

  • DoTA-RAG: Dynamic of Thought Aggregation RAG, arXiv, 2506.12571, arxiv, pdf, cication: -1

    Saksorn Ruangtanusak, Natthapath Rungseesiripak, Peerawat Rojratchadakorn, ..., Monthol Charattrakool, Natapong Nitarach

  • UniversalRAG: Retrieval-Augmented Generation over Corpora of Diverse Modalities and Granularities, arXiv, 2504.20734, arxiv, pdf, cication: -1

    Woongyeong Yeo, Kangsan Kim, Soyeong Jeong, ..., Jinheon Baek, Sung Ju Hwang ยท (universalrag.github)

  • ReasonIR: Training Retrievers for Reasoning Tasks, arXiv, 2504.20595, arxiv, pdf, cication: -1

    Rulin Shao, Rui Qiao, Varsha Kishore, ..., Pang Wei Koh, Luke Zettlemoyer

  • ๐ŸŒŸ NodeRAG: Structuring Graph-based RAG with Heterogeneous Nodes, arXiv, 2504.11544, arxiv, pdf, cication: -1

    Tianyang Xu, Haojie Zheng, Chengze Li, ..., Ruoxi Chen, Lichao Sun ยท (NodeRAG. - Terry-Xu-666) Star

  • RouterRetriever: Routing over a Mixture of Expert Embedding Models, arXiv, 2409.02685, arxiv, pdf, cication: -1

    Hyunji Lee, Luca Soldaini, Arman Cohan, ..., Minjoon Seo, Kyle Lo ยท (๐•)

  • Enhancing Financial Time-Series Forecasting with Retrieval-Augmented Large Language Models, arXiv, 2502.05878, arxiv, pdf, cication: -1

    Mengxi Xiao, Zihao Jiang, Lingfei Qian, ..., Sophia Ananiadou, Qianqian Xie ยท (huggingface)

  • SafeRAG: Benchmarking Security in Retrieval-Augmented Generation of Large Language Model, arXiv, 2501.18636, arxiv, pdf, cication: -1

    Xun Liang, Simin Niu, Zhiyu Li, ..., Mengwei Wang, Jiawei Yang ยท (SafeRAG - IAAR-Shanghai) Star

  • SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models, arXiv, 2502.09604, arxiv, pdf, cication: -1

    Yung-Sung Chuang, Benjamin Cohen-Wang, Shannon Zejiang Shen, ..., Shang-Wen Li, Wen-tau Yih

  • Chain-of-Retrieval Augmented Generation, arXiv, 2501.14342, arxiv, pdf, cication: -1

    Liang Wang, Haonan Chen, Nan Yang, ..., Zhicheng Dou, Furu Wei

  • MiniRAG: Towards Extremely Simple Retrieval-Augmented Generation, arXiv, 2501.06713, arxiv, pdf, cication: -1

    Tianyu Fan, Jingyuan Wang, Xubin Ren, ..., Chao Huang ยท (MiniRAG - HKUDS) Star

  • ๐ŸŒŸ OmniThink: Expanding Knowledge Boundaries in Machine Writing through Thinking, arXiv, 2501.09751, arxiv, pdf, cication: -1

    Zekun Xi, Wenbiao Yin, Jizhan Fang, ..., Fei Huang, Huajun Chen ยท (zjunlp.github)

  • Knowledge Models Combine Retrieval with Generation: An Introduction to RAG

  • Personalized Graph-Based Retrieval for Large Language Models, arXiv, 2501.02157, arxiv, pdf, cication: -1

    Steven Au, Cameron J. Dimacali, Ojasmitha Pedirappagari, ..., Ryan A. Rossi, Nesreen K. Ahmed

  • GeAR: Generation Augmented Retrieval, arXiv, 2501.02772, arxiv, pdf, cication: -1

    Haoyu Liu, Shaohan Huang, Jianfeng Liu, ..., Furu Wei, Qi Zhang

  • Don't Do RAG: When Cache-Augmented Generation is All You Need for Knowledge Tasks, arXiv, 2412.15605, arxiv, pdf, cication: -1

    Brian J Chan, Chao-Ting Chen, Jui-Hung Cheng, ..., Hen-Hsen Huang ยท (cag - hhhuang) Star

  • Long Context vs. RAG for LLMs: An Evaluation and Revisits, arXiv, 2501.01880, arxiv, pdf, cication: -1

    Xinze Li, Yixin Cao, Yubo Ma, ..., Aixin Sun

  • SKETCH: Structured Knowledge Enhanced Text Comprehension for Holistic Retrieval, arXiv, 2412.15443, arxiv, pdf, cication: -1

    Aakash Mahalingam, Vinesh Kumar Gande, Aman Chadha, ..., Vinija Jain, Divya Chaudhary

  • GemmaEmbed is a dense-vector embedding model, trained especially for retrieval. ๐Ÿค—

  • ๐ŸŒŸ Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference, arXiv, 2412.13663, arxiv, pdf, cication: -1

    Benjamin Warner, Antoine Chaffin, Benjamin Claviรฉ, ..., Jeremy Howard, Iacopo Poli ยท (huggingface) ยท (๐•)

  • Auto-RAG: Autonomous Retrieval-Augmented Generation for Large Language Models, arXiv, 2411.19443, arxiv, pdf, cication: -1

    Tian Yu, Shaolei Zhang, Yang Feng ยท (Auto-RAG - ictnlp) Star

  • NV-Embed-v2, a generalist embedding model that ranks No. 1 on the Massive Text Embedding Benchmark (MTEB benchmark) ๐Ÿค—

    ยท (arxiv)

  • Jina CLIP v2: Multilingual Multimodal Embeddings for Texts and Images ๐Ÿค—

    ยท (huggingface)

  • Long Term Memory: The Foundation of AI Self-Evolution, arXiv, 2410.15665, arxiv, pdf, cication: -1

    Xun Jiang, Feng Li, Han Zhao, ..., Mengdi Wang, Tianqiao Chen ยท (๐•)

  • Binary vector embeddings are so cool

  • ๐ŸŒŸ HtmlRAG: HTML is Better Than Plain Text for Modeling Retrieved Knowledge in RAG Systems, arXiv, 2411.02959, arxiv, pdf, cication: -1

    Jiejun Tan, Zhicheng Dou, Wen Wang, ..., Weipeng Chen, Ji-Rong Wen

  • M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding, arXiv, 2411.04952, arxiv, pdf, cication: -1

    Jaemin Cho, Debanjan Mahata, Ozan Irsoy, ..., Yujie He, Mohit Bansal ยท (m3docrag.github)

  • In Defense of RAG in the Era of Long-Context Language Models, arXiv, 2409.01666, arxiv, pdf, cication: 3

    Tan Yu, Anbang Xu, Rama Akkiraju ยท (zyphra)

  • Toward General Instruction-Following Alignment for Retrieval-Augmented Generation, arXiv, 2410.09584, arxiv, pdf, cication: -1

    Guanting Dong, Xiaoshuai Song, Yutao Zhu, ..., Zhicheng Dou, Ji-Rong Wen ยท (FollowRAG.github) ยท (arxiv) ยท (FollowRAG - dongguanting) Star ยท (huggingface)

  • Meta-Chunking: Learning Efficient Text Segmentation via Logical Perception, arXiv, 2410.12788, arxiv, pdf, cication: -1

    Jihao Zhao, Zhiyuan Ji, Pengnian Qi, ..., Feiyu Xiong, Zhiyu Li ยท (Meta-Chunking - IAAR-Shanghai) Star ยท (arxiv)

  • Your Mixture-of-Experts LLM Is Secretly an Embedding Model For Free, arXiv, 2410.10814, arxiv, pdf, cication: -1

    Ziyue Li, Tianyi Zhou ยท (MoE-Embedding - tianyi-lab) Star

Multi Modal

Embedding

Evaluation

  • How to Evaluate Document Extraction ๐•

  • WebWalker: Benchmarking LLMs in Web Traversal, arXiv, 2501.07572, arxiv, pdf, cication: -1

    Jialong Wu, Wenbiao Yin, Yong Jiang, ..., Pengjun Xie, Fei Huang ยท (WebWalker - Alibaba-NLP) Star

  • ๐ŸŒŸ MMDocIR: Benchmarking Multi-Modal Retrieval for Long Documents, arXiv, 2501.08828, arxiv, pdf, cication: -1

    Kuicai Dong, Yujing Chang, Xin Deik Goh, ..., Ruiming Tang, Yong Liu

  • OmniEval: An Omnidirectional and Automatic RAG Evaluation Benchmark in Financial Domain, arXiv, 2412.13018, arxiv, pdf, cication: -1

    Shuting Wang, Jiejun Tan, Zhicheng Dou, ..., Ji-Rong Wen ยท (OmniEval - RUC-NLPIR) Star

  • Long Context RAG Performance of Large Language Models, arXiv, 2411.03538, arxiv, pdf, cication: -1

    Quinn Leng, Jacob Portes, Sam Havens, ..., Matei Zaharia, Michael Carbin

  • CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation, arXiv, 2410.23090, arxiv, pdf, cication: -1

    Yiruo Cheng, Kelong Mao, Ziliang Zhao, ..., Ji-Rong Wen, Zhicheng Dou ยท (CORAL - Ariya12138) Star

Database

Projects

Products

Misc

Vector Database

What's inside

10 sections: Survey, RAG, Multi Modal, Embedding, Evaluation, Database, Projects, Products, Misc, Vector Database

Change this for your project

  • Replace metame-ai/awesome-llm-plaza with your own repository name if forking
  • Replace DavidZWZ/Awesome-RAG-Reasoning with your own survey repo link if maintaining a fork
  • Replace HKUDS/MiniRAG with your own project reference if customizing the list

Where it goes

Reference documentation for a retrieval pipeline. Keep with the ingestion or retrieval code it describes.

Worth borrowing

  • Grouping resources by sub-topic (e.g., Multi Modal, Evaluation) for quick scanning
  • Using star badges and external links (arXiv, Hugging Face, GitHub) to surface quality and provenance

Related Documents