All Documents

3,528 documents available

EVALS.md

Enhancing Budget-Aware Gating for Retrieval Augmented Generation (RAG)

Describes a thesis that designs and validates a Budget-Aware Uncertainty Gate (BAUG) to reduce token consumption and latency in multi-round RAG systems while maintaining answer quality.

aiagentllm
0
0
inasfarras
RAG.md

JudgeIt (From SuperKnowa)- Automatic Eval Framework for Gen AI Pipelines

Automates GenAI pipeline evaluation using an LLM-as-a-judge, replacing slow human evaluation for RAG, multi-turn query rewrite, and agentic workflows.

aiagentllm
0
0
ibm-self-serve-assets
ARCHITECTURE.md

Support Agent Chatbot Implementation Plan

Outlines a multi-channel customer support chatbot using Cloudflare Workers, RAG, and vector search across multiple businesses and personas.

aiagentrag
0
2
ai-primitives
RAG.md

NLP - Dialogue System

Lists 12 influential papers on multi-turn dialogue systems with conference venues and one-sentence summaries of each contribution.

aieval
0
0
zhongpeixiang
ARCHITECTURE.md

Warp AI Coding Preferences for Ragex

Defines coding conventions, architecture, and implementation phases for an Elixir-based MCP server that performs hybrid RAG codebase analysis.

airageval
0
2
Oeditus
SKILL.md

RAG Content Chunker Skill

Splits text or Markdown into token-counted chunks with deterministic IDs for RAG pipelines, supporting three chunking strategies.

aiagentrag
0
1
labrat-0
FAQ.md

❓ Frequently Asked Questions (FAQ)

Answers common questions about the AI Engagement Accelerator Kit's contents, setup, and usage for cross-functional GenAI project teams.

airag
0
0
stanchat
SPEC.md

SPEC: MCP Mode Skill Generation

Replaces MCP mode's raw document dump with an LLM-generated SKILL.md file containing actionable instructions and code patterns.

aillmrag
0
1
KasarLabs
RAG.md

llm-hallucination-survey

Curates a structured bibliography of 100+ papers on LLM hallucination, organised by evaluation, source, and mitigation.

aillmeval
0
5
HillZhang1999
RAG.md

📈 AlphaInsight Pro - Final Year Project Demo

Describes a production-deployed financial RAG demo app with PDF processing, semantic search, and AI Q&A via a Bloomberg-style terminal UI.

aillmrag
0
0
RohanExploit
RAG.md

040: Memory Retrieval Pipeline

Defines a 4-stage memory retrieval pipeline with three retrieval modes, entity synthesis, dedup, contradiction detection, supersession, and file-backed freshness scanning.

aiagentllm
0
0
samhotchkiss
RAG.md

Overview

Explains how LocusZoom.js retrieves and combines data from multiple sources using adapters, namespaces, and data operations.

aieval
0
0
statgen
RAG.md

RALM_Survey

Catalogs over 100 papers on retrieval-augmented language models, organized by definition, retriever, LM, enhancement, data source, and application.

aillmrag
0
0
2471023025
RAG.md

README

Indexes 20+ KDB.AI sample notebooks covering vector database use cases from quickstarts to multimodal RAG and time-series search.

airageval
0
0
KxSystems
RAG.md

Information Retrieval

Curates a collection of foundational and neural IR resources, including papers, courses, and tutorials, with brief annotations.

airageval
0
0
brylevkirill
RAG.md

Interactive RAG with MongoDB Atlas + Function Calling API

Explains interactive RAG with MongoDB Atlas and OpenAI function calling, enabling dynamic retrieval strategy adjustments.

aillmrag
0
0
ranfysvalle02
RAG.md

Notes

Collects RAG optimization techniques, a MicroSaaS business plan, and a LangChain development strategy into a single personal reference note.

aillmrag
0
0
yogeshhk
PROMPTS.md

FlexFlow Serve: Low-Latency, High-Performance LLM Serving

Introduces FlexFlow Serve, an open-source compiler and distributed system for low-latency, high-performance LLM serving with speculative inference.

aillm
0
2
flexflow
RAG.md

Curso de Chatbots - Implementaciones RAG

Teaches RAG through four Python implementations that load documents, generate embeddings, and answer queries using OpenAI or local models.

aillmrag
0
2
educep
RAG.md

Quickly build Generative AI applications with Amazon Bedrock

Provides runnable code samples for image generation, text tasks, chatbots, and RAG using Amazon Bedrock foundation models.

airagprompt
0
0
build-on-aws
RAG.md

FlashRAG

Explains FlashRAG's SKR method, retrieval algorithms (DPR, E5, BGE, ANCE), and multi-stage pipeline (recall, coarse ranking, fine ranking, reranking) for RAG systems.

aillmrag
0
3
1850298154
RAG.md

softrag

Provides a local-first RAG library using SQLite with sqlite-vec for document, embedding, and cache storage in a single.db file.

airagprompt
0
0
JulioPeixoto
RAG.md

Retrieval Augmented Generation

Implements a local Retrieval-Augmented Generation system using LangChain, Ollama, and Hugging Face models to augment LLM responses with context from a vector store.

aiagentllm
0
0
idra-lab
ARCHITECTURE.md

Agentic RAG In This Project

Documents the agentic RAG pipeline, covering retrieval, entity extraction, tool planning, execution, and response synthesis with trace output.

aiagentllm
0
2
hoangsonww
Page 124 of 147