Awesome AI Agent Stack
RAG & Retrieval
Vector databases, search engines, embeddings and the frameworks that ground answers in your own data.
Vector databases
- chroma-core/chroma - Embedded, zero-config vector store; the right default for prototypes.
- qdrant/qdrant - Rust vector database with filtering and hybrid search; production-ready.
- pgvector/pgvector - Vector similarity search inside Postgres; skip a new database if you already run Postgres.
- lancedb/lancedb - Serverless, file-based vector database on Lance columnar format.
- milvus-io/milvus - Distributed vector database for billion-scale workloads.
- weaviate/weaviate - Vector database with built-in vectorizers and hybrid search.
- alibaba/zvec - A lightweight, lightning-fast, in-process vector database.
- VexDB-THU/VexDB-Lite - A cross-platform vector database integrable into existing applications.
- vearch/vearch - Distributed vector search for AI-native applications.
- vdaas/vald - Vald: a highly scalable distributed vector search engine.
- dingodb/dingo - A multi-modal vector database supporting upserts and vector queries.
- endee-io/endee - A high-performance vector database built for billions of vectors.
- philippgille/chromem-go - Embeddable vector database for Go with a Chroma-like interface.
- oceanbase/seekdb - The AI-native search database for agent storage, unifying vector search.
- Stevenic/vectra - Local vector database for Node.js, Pinecone-like but file-based.
- epsilla-cloud/vectordb - A high-performance vector database leveraging parallel graph computing.
- sdan/vlite - A fast vector database built in numpy.
- Semafind/semadb - No-fuss multi-index hybrid vector database and search engine.
- nuclia/nucliadb - NucliaDB, the AI search database for RAG.
- infiniflow/infinity - AI-native database for LLM applications with vector and full-text search.
- objectbox/objectbox-java - On-device database for Android and JVM with vector search.
Vector search libraries
Search engines
Embeddings, rerankers & search
RAG frameworks & techniques
- neuml/txtai - All-in-one AI framework for semantic search and LLM orchestration.
- OSU-NLP-Group/HippoRAG - [NeurIPS'24] RAG framework inspired by human long-term memory for deeper multi-hop reasoning.
- carriex/recomp - RECOMP: compress and selectively augment retrieved documents for RAG.
- HKUDS/RAG-Anything - All-in-one multimodal RAG framework.
- NirDiamant/RAG_Techniques - Advanced RAG techniques, each with runnable code.
- Shubhamsaboo/all-rag-techniques - Every RAG technique, implemented simply.
- pguso/rag-from-scratch - Build RAG from scratch with local models, no black boxes.
- jamwithai/production-agentic-rag-course - Production agentic RAG course.
- upstash/context7 - Up-to-date library documentation served to LLMs and editors.
- pathwaycom/pathway - Python ETL framework for stream processing, real-time analytics, LLM pipelines, and RAG.
- pathwaycom/llm-app - Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data.
- The-Vibe-Company/quivr - Opiniated RAG for integrating GenAI in your apps 🧠Focus on your product rather than the RAG.
- onyx-dot-app/onyx - Open Source AI Platform - AI Chat with advanced features that works with every LLM.
- infiniflow/ragflow - RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses.
- zylon-ai/private-gpt - Complete API layer for private AI applications on local models: RAG, skills, tools, MCP.
- VectifyAI/PageIndex - PageIndex: Document Index for Vectorless, Reasoning-based RAG.
- llmware-ai/llmware - Unified framework for enterprise RAG pipelines with small specialized models.
- OpenBMB/UltraRAG - Low-code MCP framework for building complex, innovative RAG pipelines.
- Azure/GPT-RAG - Enterprise-grade accelerator for agentic RAG on Azure.
- memfreeme/memfree - MemFree: hybrid AI search engine and AI page generator.
- ggozad/haiku.rag - Agentic RAG for local and self-hosted document search with hybrid retrieval.
- notadev-iamaura/OneRAG - Production-ready RAG framework with 1-line config swaps for providers.
- MoMoM101/RAG-ReActAgent - Retrieval-augmented generation system with a ReAct agent loop.
- flamehaven01/Flamehaven-Filesearch - Self-hosted RAG search engine: 34 formats, BM25 plus hybrid search.
- hanxiao/searchbox - Airgapped closed-corpus QA: a self-hosted agent explores documents and answers.
- nshkrdotcom/rag_ex - Elixir RAG library with multi-LLM routing across Gemini, Claude, and OpenAI.
- weizhepei/InstructRAG - [ICLR 2025] InstructRAG: instructing RAG via self-synthesized rationales.
- Tencent/WeKnora - LLM knowledge platform that turns documents into queryable RAG and agent workflows.
- MODSetter/SurfSense - Privacy-focused open source NotebookLM alternative with RAG over your own sources.
- zilliztech/deep-searcher - Deep research over private data using reasoning models and vector search.
- labring/FastGPT - FastGPT is a knowledge-based platform built on the LLMs, offers a comprehensive suite of.
- Minima-AI-Inc/minima - On-premises conversational RAG with configurable containers.
- netease-youdao/QAnything - Question answering over documents with a web UI.
- SciPhi-AI/R2R - Production-ready agentic RAG system with a dashboard.
- morphik-org/morphik-core - Open-source multimodal retrieval engine for documents, images and video.
Graph RAG
Semantic caching
RAG evaluation & benchmarks