Foundation Models

Open-weight LLMs, fine-tuning stacks, and tooling around foundation models.

113 projects

See methodology for ranking rules; order uses public GitHub metrics within this scenario.

81–100 of 113

Rank Project Stars Forks
81 open_flamingo

An open-source framework for training large multimodal models.

4.1K 319
82 OpenCoder-llm

The Open Cookbook for Top-Tier Code Large Language Model

2.1K 127
83 Aria

Codebase for Aria - an Open Multimodal Native MoE

1.1K 89
84 DeepSeek-LLM

DeepSeek LLM: Let there be answers

7.3K 1.3K
85 magicoder

[ICML'24] Magicoder: Empowering Code Generation with OSS-Instruct

2.1K 170
86 dolly

Databricks’ Dolly, a large language model trained on the Databricks Machine Learning Platform

10.8K 1.1K
87 LLaMA-Adapter

[ICLR 2024] Fine-tuning LLaMA to follow Instructions within 1 Hour and 1.2M Parameters

5.9K 378
88 XrayGLM

The first Chinese Medical Multimodal Model that Chest Radiographs Summarization.

1.1K 144
89 DeepSeek-VL

DeepSeek-VL: Towards Real-World Vision-Language Understanding

4.2K 597
90 bpemb

Pre-trained subword embeddings in 275 languages, based on Byte-Pair Encoding (BPE)

1.2K 100
91 ONE-PEACE

A general representation model across vision, audio, language modalities. Paper: ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

1.1K 71
92 torchscale

Foundation Architecture for (M)LLMs

3.1K 225
93 Linly

Chinese-LLaMA 1&2、Chinese-Falcon 基础模型;ChatFlow中文对话模型;中文OpenLLaMA模型;NLP预训练/指令微调数据集

3.0K 222
94 OFA

Official repository of OFA (ICML 2022). Paper: OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Learning Framework

2.6K 248
95 Otter

🦦 Otter, a multi-modal model based on OpenFlamingo (open-sourced version of DeepMind's Flamingo), trained on MIMIC-IT and showcasing improved instruction-following and in-context learning ability.

3.4K 210
96 open_llama

OpenLLaMA, a permissively licensed open source reproduction of Meta AI’s LLaMA 7B trained on the RedPajama dataset

7.5K 404
97 VisCPM

[ICLR'24 spotlight] Chinese and English Multimodal Large Model Series (Chat and Paint)

1.1K 88
98 ChatLM-mini-Chinese

中文对话0.2B小模型(ChatLM-Chinese-0.2B),开源所有数据集来源、数据清洗、tokenizer训练、模型预训练、SFT指令微调、RLHF优化等流程的全部代码。支持下游任务sft微调,给出三元组信息抽取微调示例。

1.7K 192
99 SPIN

The official implementation of Self-Play Fine-Tuning (SPIN)

1.3K 106
100 AliceMind

ALIbaba's Collection of Encoder-decoders from MinD (Machine IntelligeNce of Damo) Lab

2.0K 301