GitHub repo leaderboard by stars, growth rate and activity.
RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs
Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, PostgreSQL, real-time data APIs, and more.
[EMNLP2025] LightRAG: Simple and Fast Retrieval-Augmented Generation
Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and Llama) RAG and Agent app with langchain
📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG
An LLM-powered knowledge curation system that researches a topic and generates a full-length report with citations.
This repository showcases various advanced techniques for Retrieval-Augmented Generation (RAG) systems. Each technique has a detailed notebook tutorial.
Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over retrieval, routing, memory, and generation. Built for scalable agents, RAG, multimodal applications, semantic search, and conversational systems.
Memory layer for AI Agents. Replace complex RAG pipelines with a serverless, single-file memory layer. Give your agents instant retrieval and long-term memory.
Unified framework for building enterprise RAG pipelines with small, specialized models
| # | Repo | Language | Stars | 30-day trend | Last updated |
|---|---|---|---|---|---|
| 1 | RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs | Go | 90,410 | last pushed 24 hours ago | |
| 2 | Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, PostgreSQL, real-time data APIs, and more. | Jupyter Notebook | 58,943 | last pushed 2 months ago | |
| 3 | [EMNLP2025] LightRAG: Simple and Fast Retrieval-Augmented Generation | Python | 39,526 | last pushed Yesterday | |
| 4 | Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and Llama) RAG and Agent app with langchain | Python | 38,631 | last pushed 10 months ago | |
| 5 | 📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG | Python | 35,604 | last pushed 23 hours ago | |
| 6 | An LLM-powered knowledge curation system that researches a topic and generates a full-length report with citations. | Python | 31,258 | last pushed 12 months ago | |
| 7 | This repository showcases various advanced techniques for Retrieval-Augmented Generation (RAG) systems. Each technique has a detailed notebook tutorial. | Jupyter Notebook | 29,424 | last pushed 6 days ago | |
| 8 | Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over retrieval, routing, memory, and generation. Built for scalable agents, RAG, multimodal applications, semantic search, and conversational systems. | Python | 26,461 | last pushed 2 days ago | |
| 9 | Memory layer for AI Agents. Replace complex RAG pipelines with a serverless, single-file memory layer. Give your agents instant retrieval and long-term memory. | Rust | 16,532 | last pushed 2 months ago | |
| 10 | Unified framework for building enterprise RAG pipelines with small, specialized models | Python | 14,848 | last pushed 4 months ago |
All · 11,474