GitHub repo leaderboard by stars, growth rate and activity.
Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search
Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and Llama) RAG and Agent app with langchain
A vector index built on TurboQuant, written in Rust with Python bindings
Memory layer for AI Agents. Replace complex RAG pipelines with a serverless, single-file memory layer. Give your agents instant retrieval and long-term memory.
A lightweight, lightning-fast, in-process vector database
[MLsys2026 Best Paper]: https://arxiv.org/abs/2506.08276. RAG on Everything with LEANN. Enjoy 97% storage savings while running a fast, accurate, and 100% private RAG application on your personal device.
AdalFlow: The library to build & auto-optimize LLM applications.
This repository contains various advanced techniques for Retrieval-Augmented Generation (RAG) systems.
Text-To-Speech, RAG, and LLMs. All local!
ChatWeb can crawl web pages, read PDF, DOCX, TXT, and extract the main content, then answer your questions based on the content, or summarize the key points.
| # | Repo | Language | Stars | 30-day trend | Last updated |
|---|---|---|---|---|---|
| 1 | Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search | Go | 46,038 | last pushed 23 hours ago | |
| 2 | Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and Llama) RAG and Agent app with langchain | Python | 38,631 | last pushed 10 months ago | |
| 3 | A vector index built on TurboQuant, written in Rust with Python bindings | Rust | 16,734 | last pushed 3 weeks ago | |
| 4 | Memory layer for AI Agents. Replace complex RAG pipelines with a serverless, single-file memory layer. Give your agents instant retrieval and long-term memory. | Rust | 16,532 | last pushed 2 months ago | |
| 5 | A lightweight, lightning-fast, in-process vector database | C++ | 15,871 | last pushed 24 hours ago | |
| 6 | [MLsys2026 Best Paper]: https://arxiv.org/abs/2506.08276. RAG on Everything with LEANN. Enjoy 97% storage savings while running a fast, accurate, and 100% private RAG application on your personal device. | Python | 12,926 | last pushed 6 days ago | |
| 7 | AdalFlow: The library to build & auto-optimize LLM applications. | Python | 4,213 | last pushed 3 months ago | |
| 8 | This repository contains various advanced techniques for Retrieval-Augmented Generation (RAG) systems. | Jupyter Notebook | 2,572 | last pushed 2 years ago | |
| 9 | Text-To-Speech, RAG, and LLMs. All local! | JavaScript | 1,912 | last pushed 2 years ago | |
| 10 | ChatWeb can crawl web pages, read PDF, DOCX, TXT, and extract the main content, then answer your questions based on the content, or summarize the key points. | Python | 916 | last pushed 4 months ago |
All · 11,474