#182 · Primary category: MLOps & Evaluation
Leaderboard
SpeechIO Leaderboard: a large, robust, comprehensive, benchmarking platform for Automatic Speech Recognition.
Project last updated:03/29/25
GitHub Stars
550
Forks
73
Contributors
19
License
Other
Why we included this project
Speech recognition is only as trustworthy as the data you test it on, and this project makes that testing unusually easy. It bundles a large set of test corpora in English and Chinese, from academic standards like LibriSpeech and AISHELL to messier real-world material such as live broadcasts, podcasts, and TV interviews, each tagged with a difficulty rating. Alongside the data sits a model zoo of commercial APIs and open-source recognizers, plus a pipeline that handles data preparation, recognition, post-processing, and error-rate scoring. The result is that you can reproduce published results, run your own model against a wide range of systems on identical data, and see how a recognizer holds up outside clean academic conditions. For anyone choosing or tuning an ASR engine, the professionally transcribed test sets are a good way to surface weaknesses that standard benchmarks miss.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
unsloth
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, FLUX and more.
LlamaFactory
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
airflow
Apache Airflow - A platform to programmatically author, schedule, and monitor workflows
langfuse
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23
netron
Visualizer for neural network, deep learning and machine learning models