#72 · Primary category: Foundation Models
Taiwan-LLM
Traditional Mandarin LLMs for Taiwan
Project last updated:04/20/25
GitHub Stars
1.4K
Forks
119
Contributors
12
License
Apache-2.0
Why we included this project
Chinese LLMs are usually built around Simplified Chinese and mainland conventions, which leaves Taiwan's Traditional Mandarin speakers underserved. Taiwan-LLM is a set of models fine-tuned on Traditional Mandarin and English data with Taiwan's usage patterns and cultural context in mind, including the 70B-parameter Llama-3-Taiwan-70B, which delivers state-of-the-art results on Traditional Mandarin NLP benchmarks. The repo doubles as a working toolkit: it ships Axolotl fine-tuning configs, links to hosted demos and model weights, and documents partnerships with organizations in medical, manufacturing, and legal fields, so teams can adapt the models to their own domain instead of starting from a generic base. It is research-grade software, so expect to benchmark it against your own data before relying on it in production.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
transformers
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
CLIP
CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
MiniCPM-V
A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone
generative-models
Generative Models by Stability AI
unilm
Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities