GitHub repo leaderboard by stars, growth rate and activity.
Making large AI models cheaper, faster and more accessible
[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
Janus-Series: Unified Multimodal Understanding and Generation Models
OpenSquilla — Token-Efficient AI Agent with same budget, higher intelligence density
Code and models for ICML 2024 paper, NExT-GPT: Any-to-Any Multimodal Large Language Model
🦦 Otter, a multi-modal model based on OpenFlamingo (open-sourced version of DeepMind's Flamingo), trained on MIMIC-IT and showcasing improved instruction-following and in-context learning ability.
[CVPR2024 Highlight][VideoChatGPT] ChatGPT with video understanding! And many more supported LMs such as miniGPT4, StableLM, and MOSS.
SuperCLUE: 中文通用大模型综合性基准 | A Benchmark for Foundation Models in Chinese
A practical lab for building, testing, and evaluating apps with Apple's Foundation Models framework.
| # | Repo | Language | Stars | 30-day trend | Last updated |
|---|---|---|---|---|---|
| 1 | Making large AI models cheaper, faster and more accessible | Python | 41,441 | last pushed 1 week ago | |
| 2 | [NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond. | Python | 25,016 | last pushed 2 years ago | |
| 3 | Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities | Python | 22,205 | last pushed 2 weeks ago | |
| 4 | Janus-Series: Unified Multimodal Understanding and Generation Models | Python | 17,762 | last pushed 2 years ago | |
| 5 | OpenSquilla — Token-Efficient AI Agent with same budget, higher intelligence density | Python | 7,000 | last pushed Yesterday | |
| 6 | Code and models for ICML 2024 paper, NExT-GPT: Any-to-Any Multimodal Large Language Model | Python | 3,634 | last pushed 1 year ago | |
| 7 | 🦦 Otter, a multi-modal model based on OpenFlamingo (open-sourced version of DeepMind's Flamingo), trained on MIMIC-IT and showcasing improved instruction-following and in-context learning ability. | Python | 3,438 | last pushed 3 years ago | |
| 8 | [CVPR2024 Highlight][VideoChatGPT] ChatGPT with video understanding! And many more supported LMs such as miniGPT4, StableLM, and MOSS. | Python | 3,349 | last pushed 2 months ago | |
| 9 | SuperCLUE: 中文通用大模型综合性基准 | A Benchmark for Foundation Models in Chinese | — | 3,298 | last pushed 7 months ago | |
| 10 | A practical lab for building, testing, and evaluating apps with Apple's Foundation Models framework. | Swift | 1,178 | last pushed 2 weeks ago |
All · 11,474