GitHub repo leaderboard by stars, growth rate and activity.
Burn is a next generation tensor library and Deep Learning Framework that doesn't compromise on flexibility, efficiency and portability.
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk
A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.
A nearly-live implementation of OpenAI's Whisper.
| # | Repo | Language | Stars | 30-day trend | Last updated |
|---|---|---|---|---|---|
| 1 | Burn is a next generation tensor library and Deep Learning Framework that doesn't compromise on flexibility, efficiency and portability. | Rust | 15,889 | last pushed Yesterday | |
| 2 | LMCache: Supercharge Your LLM with the Fastest KV Cache Layer | Python | 11,733 | last pushed 23 hours ago | |
| 3 | Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk | C++ | 5,674 | last pushed 23 hours ago | |
| 4 | A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances. | Python | 5,647 | last pushed 22 hours ago | |
| 5 | A nearly-live implementation of OpenAI's Whisper. | Python | 4,256 | last pushed 4 days ago |
All · 11,474