Foundation Models
Open-weight LLMs, fine-tuning stacks, and tooling around foundation models.
113 projects
See methodology for ranking rules; order uses public GitHub metrics within this scenario.
| Rank | Project | Stars | Forks | Updated | License |
|---|---|---|---|---|---|
| 61 |
MiniMax-01
The official repo of MiniMax-Text-01 and MiniMax-VL-01, large-language-model & vision-language-model based on Linear Attention |
3.5K | 332 | 07/07/25 | MIT |
| 62 |
MiniMax-M1
MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model. |
3.2K | 283 | 07/07/25 | Apache-2.0 |
| 63 |
DeepSeek-VL2
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding |
5.4K | 1.8K | 02/26/25 | MIT |
| 64 |
LWM
Large World Model -- Modeling Text and Video with Millions Context |
7.4K | 563 | 10/19/24 | Apache-2.0 |
| 65 |
Chinese-Vicuna
Chinese-Vicuna: A Chinese Instruction-following LLaMA-based Model —— 一个中文低资源的llama+lora方案,结构参考alpaca |
4.1K | 405 | 04/18/25 | Apache-2.0 |
| 66 |
NExT-GPT
Code and models for ICML 2024 paper, NExT-GPT: Any-to-Any Multimodal Large Language Model |
3.6K | 360 | 05/13/25 | BSD-3-Clause |
| 67 |
InternLM-XComposer
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions |
2.9K | 175 | 05/26/25 | Apache-2.0 |
| 68 |
mPLUG-DocOwl
mPLUG-DocOwl: Modularized Multimodal Large Language Model for Document Understanding |
2.4K | 154 | 05/30/25 | Apache-2.0 |
| 69 |
KoBERT
Korean BERT pre-trained cased (KoBERT) |
1.4K | 377 | 06/14/25 | Apache-2.0 |
| 70 |
mPLUG-Owl
mPLUG-Owl: The Powerful Multi-modal Large Language Model Family |
2.5K | 189 | 04/02/25 | MIT |
| 71 |
openchat
OpenChat: Advancing Open-source Language Models with Imperfect Data |
5.5K | 429 | 09/13/24 | Apache-2.0 |
| 72 |
Taiwan-LLM
Traditional Mandarin LLMs for Taiwan |
1.4K | 119 | 04/20/25 | Apache-2.0 |
| 73 |
DeepSeek-V2
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model |
5.0K | 550 | 09/25/24 | MIT |
| 74 |
Baichuan2
A series of large language models developed by Baichuan Intelligent Technology |
4.1K | 294 | 11/08/24 | Apache-2.0 |
| 75 |
Lumina-T2X
Lumina-T2X is a unified framework for Text to Any Modality Generation |
2.2K | 98 | 02/16/25 | MIT |
| 76 |
large_concept_model
Large Concept Models: Language modeling in a sentence representation space |
2.4K | 213 | 01/29/25 | MIT |
| 77 |
chatglm_finetuning
chatglm 6b finetuning and alpaca finetuning |
1.5K | 170 | 03/09/25 | Apache-2.0 |
| 78 |
Baichuan-7B
A large-scale 7B pretraining language model developed by BaiChuan-Inc. |
5.6K | 501 | 07/18/24 | Apache-2.0 |
| 79 |
Skywork
Skywork series models are pre-trained on 3.2TB of high-quality multilingual (mainly Chinese and English) and code data. We have open-sourced the model, training data, evaluation data, evaluation methods, etc. |
1.5K | 148 | 03/07/25 | Other |
| 80 |
TigerBot
TigerBot: A multi-language multi-task LLM |
2.3K | 188 | 12/28/24 | Apache-2.0 |