Foundation Models
Open-weight LLMs, fine-tuning stacks, and tooling around foundation models.
113 projects
See methodology for ranking rules; order uses public GitHub metrics within this scenario.
| Rank | Project | Stars | Forks | Updated | License |
|---|---|---|---|---|---|
| 81 |
open_flamingo
An open-source framework for training large multimodal models. |
4.1K | 319 | 08/31/24 | MIT |
| 82 |
OpenCoder-llm
The Open Cookbook for Top-Tier Code Large Language Model |
2.1K | 127 | 12/08/24 | MIT |
| 83 |
Aria
Codebase for Aria - an Open Multimodal Native MoE |
1.1K | 89 | 01/22/25 | Apache-2.0 |
| 84 |
DeepSeek-LLM
DeepSeek LLM: Let there be answers |
7.3K | 1.3K | 02/04/24 | MIT |
| 85 |
magicoder
[ICML'24] Magicoder: Empowering Code Generation with OSS-Instruct |
2.1K | 170 | 11/01/24 | MIT |
| 86 |
dolly
Databricks’ Dolly, a large language model trained on the Databricks Machine Learning Platform |
10.8K | 1.1K | 06/30/23 | Apache-2.0 |
| 87 |
LLaMA-Adapter
[ICLR 2024] Fine-tuning LLaMA to follow Instructions within 1 Hour and 1.2M Parameters |
5.9K | 378 | 03/14/24 | GPL-3.0 |
| 88 |
XrayGLM
The first Chinese Medical Multimodal Model that Chest Radiographs Summarization. |
1.1K | 144 | 11/20/24 | Other |
| 89 |
DeepSeek-VL
DeepSeek-VL: Towards Real-World Vision-Language Understanding |
4.2K | 597 | 04/24/24 | MIT |
| 90 |
bpemb
Pre-trained subword embeddings in 275 languages, based on Byte-Pair Encoding (BPE) |
1.2K | 100 | 10/01/24 | MIT |
| 91 |
ONE-PEACE
A general representation model across vision, audio, language modalities. Paper: ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities |
1.1K | 71 | 10/06/24 | Apache-2.0 |
| 92 |
torchscale
Foundation Architecture for (M)LLMs |
3.1K | 225 | 04/11/24 | MIT |
| 93 |
Linly
Chinese-LLaMA 1&2、Chinese-Falcon 基础模型;ChatFlow中文对话模型;中文OpenLLaMA模型;NLP预训练/指令微调数据集 |
3.0K | 222 | 04/14/24 | Other |
| 94 |
OFA
Official repository of OFA (ICML 2022). Paper: OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Learning Framework |
2.6K | 248 | 04/24/24 | Apache-2.0 |
| 95 |
Otter
🦦 Otter, a multi-modal model based on OpenFlamingo (open-sourced version of DeepMind's Flamingo), trained on MIMIC-IT and showcasing improved instruction-following and in-context learning ability. |
3.4K | 210 | 03/05/24 | MIT |
| 96 |
open_llama
OpenLLaMA, a permissively licensed open source reproduction of Meta AI’s LLaMA 7B trained on the RedPajama dataset |
7.5K | 404 | 07/16/23 | Apache-2.0 |
| 97 |
VisCPM
[ICLR'24 spotlight] Chinese and English Multimodal Large Model Series (Chat and Paint) |
1.1K | 88 | 06/13/24 | Other |
| 98 |
ChatLM-mini-Chinese
中文对话0.2B小模型(ChatLM-Chinese-0.2B),开源所有数据集来源、数据清洗、tokenizer训练、模型预训练、SFT指令微调、RLHF优化等流程的全部代码。支持下游任务sft微调,给出三元组信息抽取微调示例。 |
1.7K | 192 | 04/20/24 | Apache-2.0 |
| 99 |
SPIN
The official implementation of Self-Play Fine-Tuning (SPIN) |
1.3K | 106 | 05/08/24 | Apache-2.0 |
| 100 |
AliceMind
ALIbaba's Collection of Encoder-decoders from MinD (Machine IntelligeNce of Damo) Lab |
2.0K | 301 | 03/19/24 | Apache-2.0 |