Foundation Models

Open-weight LLMs, fine-tuning stacks, and tooling around foundation models.

113 projects

See methodology for ranking rules; order uses public GitHub metrics within this scenario.

101–113 of 113

Rank Project Stars Forks
101 xlnet

XLNet: Generalized Autoregressive Pretraining for Language Understanding

6.2K 1.1K
102 kogpt

KakaoBrain KoGPT (Korean Generative Pre-trained Transformer)

1.0K 135
103 MetaTransformer

Meta-Transformer for Unified Multimodal Learning

1.6K 116
104 unit-minions

AI Development Efficiency: Hands-on LoRA Training, covering Llama (Alpaca LoRA) and ChatGLM (ChatGLM Tuning) LoRA. Training tasks: user story generation, test code generation, code assistance, text-to-SQL, text-to-code, etc.

1.1K 124
105 Chinese-Llama-2-7b

First downloadable and runnable Chinese LLaMA2 model in the open-source community!

2.2K 196
106 Baichuan-13B

A 13B large language model developed by Baichuan Intelligent Technology

2.9K 229
107 CoCa-pytorch

Implementation of CoCa, Contrastive Captioners are Image-Text Foundation Models, in Pytorch

1.2K 89
108 long_llama

LongLLaMA is a large language model capable of handling long contexts. It is based on OpenLLaMA and fine-tuned with the Focused Transformer (FoT) method.

1.5K 84
109 gpt2-ml

GPT2 for Multiple Languages, including pretrained models. GPT2 多语言支持, 15亿参数中文预训练模型

1.7K 325
110 ru-gpts

Russian GPT3 models.

2.1K 433
111 scibert

A BERT model for scientific text.

1.7K 232
112 pytorch-openai-transformer-lm

🐥A PyTorch implementation of OpenAI's finetuned transformer language model with a script to import the weights pre-trained by OpenAI

1.5K 285
113 gpt-2-Pytorch

Simple Text-Generator with OpenAI gpt-2 Pytorch Implementation

1.0K 230
< Previous
/ 6
Next >