#132 · Primary category: MLOps & Evaluation
deepchecks
Deepchecks: Tests for Continuous Validation of ML Models & Data. Deepchecks is a holistic open-source solution for all of your AI & ML validation needs, enabling to thoroughly test your data and models from research to production.
Project last updated:12/28/25
GitHub Stars
4.0K
Forks
304
Contributors
56
License
Other
Why we included this project
Deepchecks is for teams that have watched a model degrade in production and wish they had caught it earlier. It packages validation as a Python library of tests for data quality, data drift, model performance, and label issues, run directly against your pandas dataframes and trained models. You can use it interactively in notebooks, generate HTML reports to share with stakeholders, or wire the same suites into CI so regressions surface before release. Rather than covering one slice of the lifecycle, it spans research through deployment, which makes it useful both for data scientists who want ready-made checks and MLOps engineers who need a consistent, documented way to gate releases.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
unsloth
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, FLUX and more.
LlamaFactory
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
airflow
Apache Airflow - A platform to programmatically author, schedule, and monitor workflows
langfuse
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23
netron
Visualizer for neural network, deep learning and machine learning models