Foundation Models
Open-weight LLMs, fine-tuning stacks, and tooling around foundation models.
113 projects
See methodology for ranking rules; order uses public GitHub metrics within this scenario.
| Rank | Project | Stars | Forks | Updated | License |
|---|---|---|---|---|---|
| 21 |
VLM-R1
Solve Visual Understanding with Reinforced VLMs |
6.0K | 384 | 07/07/26 | Apache-2.0 |
| 22 |
Chinese-LLaMA-Alpaca-2
中文LLaMA-2 & Alpaca-2大模型二期项目 + 64K超长上下文模型 (Chinese LLaMA-2 & Alpaca-2 LLMs with 64K long context models) |
7.1K | 561 | 04/19/26 | Apache-2.0 |
| 23 |
InternVL
[CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o. |
10.1K | 793 | 09/22/25 | MIT |
| 24 |
Huatuo-Llama-Med-Chinese
Repo for BenCao [original name: HuaTuo (华驼)], Instruction-tuning Large Language Models with Chinese Medical Knowledge. |
5.0K | 498 | 07/04/26 | Apache-2.0 |
| 25 |
CodeGen
CodeGen is a family of open-source model for program synthesis. Trained on TPU-v4. Competitive with OpenAI Codex. |
5.2K | 422 | 06/02/26 | Apache-2.0 |
| 26 |
ChatGLM3
ChatGLM3 series: Open Bilingual Chat LLMs |
13.7K | 1.6K | 01/13/25 | Apache-2.0 |
| 27 |
open-llms
📋 A list of open LLMs available for commercial use. |
12.9K | 987 | 02/13/25 | Apache-2.0 |
| 28 |
Fengshenbang-LM
Fengshenbang-LM is an open-source large model system by IDEA Research, serving as infrastructure for Chinese AIGC and cognitive AI. |
4.1K | 374 | 06/08/26 | Apache-2.0 |
| 29 |
LimiX
LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence https://arxiv.org/abs/2509.03505 |
4.1K | 305 | 06/16/26 | Apache-2.0 |
| 30 |
Skywork-R1V
Skywork-R1V is an advanced multimodal AI model series developed by Skywork AI, specializing in vision-language reasoning. |
3.2K | 283 | 07/29/26 | MIT |
| 31 |
Eagle
Eagle: Frontier Vision-Language Models with Data-Centric Strategies |
3.5K | 335 | 06/24/26 | Apache-2.0 |
| 32 |
InternLM
Official release of InternLM series (InternLM, InternLM2, InternLM2.5, InternLM3). |
7.3K | 509 | 10/30/25 | Apache-2.0 |
| 33 |
minimind-o
🎙️ A 0.1B Omni model trained from scratch, capable of listening, speaking, and seeing! |
2.4K | 282 | 08/06/26 | Apache-2.0 |
| 34 |
dllm
dLLM: Simple Diffusion Language Modeling |
2.7K | 280 | 07/17/26 | Apache-2.0 |
| 35 |
DeepSeek-Coder-V2
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence |
7.0K | 1.1K | 11/11/25 | MIT |
| 36 |
Qwen3-Omni
Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time. |
4.0K | 291 | 04/23/26 | Apache-2.0 |
| 37 |
WizardLM
LLMs build upon Evol Insturct: WizardLM, WizardCoder, WizardMath |
9.5K | 743 | 06/07/25 | Other |
| 38 |
ChatGLM2-6B
ChatGLM2-6B: An Open Bilingual Chat LLM |
15.5K | 1.8K | 06/27/24 | Other |
| 39 |
transfusion-pytorch
Pytorch implementation of Transfusion, "Predict the Next Token and Diffuse Images with One Multi-Modal Model", from MetaAI |
1.4K | 75 | 08/27/26 | MIT |
| 40 |
VibeThinker
Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B |
1.6K | 117 | 08/14/26 | MIT |