Foundation Models

Open-weight LLMs, fine-tuning stacks, and tooling around foundation models.

113 projects

See methodology for ranking rules; order uses public GitHub metrics within this scenario.

61–80 of 113

Rank Project Stars Forks
61 MiniMax-01

The official repo of MiniMax-Text-01 and MiniMax-VL-01, large-language-model & vision-language-model based on Linear Attention

3.5K 332
62 MiniMax-M1

MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model.

3.2K 283
63 DeepSeek-VL2

DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

5.4K 1.8K
64 LWM

Large World Model -- Modeling Text and Video with Millions Context

7.4K 563
65 Chinese-Vicuna

Chinese-Vicuna: A Chinese Instruction-following LLaMA-based Model —— 一个中文低资源的llama+lora方案,结构参考alpaca

4.1K 405
66 NExT-GPT

Code and models for ICML 2024 paper, NExT-GPT: Any-to-Any Multimodal Large Language Model

3.6K 360
67 InternLM-XComposer

InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions

2.9K 175
68 mPLUG-DocOwl

mPLUG-DocOwl: Modularized Multimodal Large Language Model for Document Understanding

2.4K 154
69 KoBERT

Korean BERT pre-trained cased (KoBERT)

1.4K 377
70 mPLUG-Owl

mPLUG-Owl: The Powerful Multi-modal Large Language Model Family

2.5K 189
71 openchat

OpenChat: Advancing Open-source Language Models with Imperfect Data

5.5K 429
72 Taiwan-LLM

Traditional Mandarin LLMs for Taiwan

1.4K 119
73 DeepSeek-V2

DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

5.0K 550
74 Baichuan2

A series of large language models developed by Baichuan Intelligent Technology

4.1K 294
75 Lumina-T2X

Lumina-T2X is a unified framework for Text to Any Modality Generation

2.2K 98
76 large_concept_model

Large Concept Models: Language modeling in a sentence representation space

2.4K 213
77 chatglm_finetuning

chatglm 6b finetuning and alpaca finetuning

1.5K 170
78 Baichuan-7B

A large-scale 7B pretraining language model developed by BaiChuan-Inc.

5.6K 501
79 Skywork

Skywork series models are pre-trained on 3.2TB of high-quality multilingual (mainly Chinese and English) and code data. We have open-sourced the model, training data, evaluation data, evaluation methods, etc.

1.5K 148
80 TigerBot

TigerBot: A multi-language multi-task LLM

2.3K 188