GitHub repo leaderboard by stars, growth rate and activity.
Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
🤯 LobeHub is your Chief Agent Operator, organizing your agents into 7×24 operations by hiring, scheduling, and reporting on your entire AI team.
The unified workspace where open-source models get things done for you.
SGLang is a high-performance serving framework for large language models and multimodal models.
FreeToken brings datacenter-scale model serving to your desktop. Run massive models locally, fast and efficiently.
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.
Command Code AI — the best coding agent for open models.
TokenSpeed is a speed-of-light LLM inference engine.
| # | Repo | Language | Stars | 30-day trend | Last updated |
|---|---|---|---|---|---|
| 1 | Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models. | Go | 180,490 | — | last pushed 2 days ago |
| 2 | 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training. | Python | 165,014 | — | last pushed 2 days ago |
| 3 | 🤯 LobeHub is your Chief Agent Operator, organizing your agents into 7×24 operations by hiring, scheduling, and reporting on your entire AI team. | TypeScript | 82,334 | — | last pushed 2 days ago |
| 4 | The unified workspace where open-source models get things done for you. | Makefile | 39,716 | — | last pushed 2 days ago |
| 5 | SGLang is a high-performance serving framework for large language models and multimodal models. | Python | 35,666 | — | last pushed 2 days ago |
| 6 | FreeToken brings datacenter-scale model serving to your desktop. Run massive models locally, fast and efficiently. | Python | 12,173 | — | last pushed 2 days ago |
| 7 | Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API. | Python | 9,555 | — | last pushed 3 days ago |
| 8 | Command Code AI — the best coding agent for open models. | — | 3,900 | — | last pushed 4 weeks ago |
| 9 | 和国产大模型相关的免费工具集 | JavaScript | 2,523 | — | last pushed 7 months ago |
| 10 | TokenSpeed is a speed-of-light LLM inference engine. | Python | 2,105 | — | last pushed 2 days ago |
All · 11,474