GitHub repo leaderboard by stars, growth rate and activity.
利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant into a full video production studio.
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
A generative speech model for daily dialogue.
Instant voice cloning by MIT and MyShell. Audio foundation model.
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
| # | Repo | Language | Stars | 30-day trend | Last updated |
|---|---|---|---|---|---|
| 1 | 利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow. | Python | 121,644 | — | last pushed 2 days ago |
| 2 | Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more. | Python | 75,889 | — | last pushed 2 days ago |
| 3 | 1 min voice data can also be used to train a good TTS model! (few shot voice cloning) | Python | 61,645 | — | last pushed 3 weeks ago |
| 4 | World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant into a full video production studio. | Python | 56,709 | — | last pushed 5 days ago |
| 5 | 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production | Python | 45,992 | — | last pushed 2 years ago |
| 6 | A generative speech model for daily dialogue. | Python | 39,822 | — | last pushed 5 months ago |
| 7 | Instant voice cloning by MIT and MyShell. Audio foundation model. | Python | 37,483 | — | last pushed 1 year ago |
| 8 | 🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time | Python | 36,910 | — | last pushed 6 months ago |
| 9 | VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning | Python | 36,862 | — | last pushed 1 week ago |
| 10 | Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability. | Python | 23,519 | — | last pushed 4 months ago |
All · 11,474