#117 · Primary category: MLOps & Evaluation
CLUE
中文语言理解测评基准 Chinese Language Understanding Evaluation Benchmark: datasets, baselines, pre-trained models, corpus and leaderboard
Project last updated:02/06/26
GitHub Stars
4.3K
Forks
544
Contributors
24
License
Other
Why we included this project
CLUE is the benchmark most people turn to when they need a fixed, comparable footing for Chinese-language NLP models. It bundles a broad set of Chinese understanding tasks, from text classification and sentence-pair similarity to natural language inference, and maintains a public leaderboard so you can see how a candidate model performs before committing to it. The repo also includes baseline implementations and links to pre-trained Chinese models, which gives you a sensible starting point instead of training from scratch. It is a research-community standard, not a product you deploy, but for anyone selecting or sanity-checking Chinese models, the datasets and leaderboard history save real time. Because it spans older and current task families, it also works as a window into how Chinese NLU has progressed.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
unsloth
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, FLUX and more.
LlamaFactory
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
airflow
Apache Airflow - A platform to programmatically author, schedule, and monitor workflows
langfuse
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23
netron
Visualizer for neural network, deep learning and machine learning models