Foundation Models

Open-weight LLMs, fine-tuning stacks, and tooling around foundation models.

113 projects

See methodology for ranking rules; order uses public GitHub metrics within this scenario.

21–40 of 113

Rank Project Stars Forks
21 VLM-R1

Solve Visual Understanding with Reinforced VLMs

6.0K 384
22 Chinese-LLaMA-Alpaca-2

中文LLaMA-2 & Alpaca-2大模型二期项目 + 64K超长上下文模型 (Chinese LLaMA-2 & Alpaca-2 LLMs with 64K long context models)

7.1K 561
23 InternVL

[CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o.

10.1K 793
24 Huatuo-Llama-Med-Chinese

Repo for BenCao [original name: HuaTuo (华驼)], Instruction-tuning Large Language Models with Chinese Medical Knowledge.

5.0K 498
25 CodeGen

CodeGen is a family of open-source model for program synthesis. Trained on TPU-v4. Competitive with OpenAI Codex.

5.2K 422
26 ChatGLM3

ChatGLM3 series: Open Bilingual Chat LLMs

13.7K 1.6K
27 open-llms

📋 A list of open LLMs available for commercial use.

12.9K 987
28 Fengshenbang-LM

Fengshenbang-LM is an open-source large model system by IDEA Research, serving as infrastructure for Chinese AIGC and cognitive AI.

4.1K 374
29 LimiX

LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence https://arxiv.org/abs/2509.03505

4.1K 305
30 Skywork-R1V

Skywork-R1V is an advanced multimodal AI model series developed by Skywork AI, specializing in vision-language reasoning.

3.2K 283
31 Eagle

Eagle: Frontier Vision-Language Models with Data-Centric Strategies

3.5K 335
32 InternLM

Official release of InternLM series (InternLM, InternLM2, InternLM2.5, InternLM3).

7.3K 509
33 minimind-o

🎙️ A 0.1B Omni model trained from scratch, capable of listening, speaking, and seeing!

2.4K 282
34 dllm

dLLM: Simple Diffusion Language Modeling

2.7K 280
35 DeepSeek-Coder-V2

DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence

7.0K 1.1K
36 Qwen3-Omni

Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.

4.0K 291
37 WizardLM

LLMs build upon Evol Insturct: WizardLM, WizardCoder, WizardMath

9.5K 743
38 ChatGLM2-6B

ChatGLM2-6B: An Open Bilingual Chat LLM

15.5K 1.8K
39 transfusion-pytorch

Pytorch implementation of Transfusion, "Predict the Next Token and Diffuse Images with One Multi-Modal Model", from MetaAI

1.4K 75
40 VibeThinker

Tiny Model, Big Logic: Diversity-Driven Optimization Elicits Large-Model Reasoning Ability in VibeThinker-1.5B

1.6K 117