#35 · Primary category: MLOps & Evaluation
iFixAi
Independent Auditing of AI Agents. Run by human or the agent itself, to answer the most crucial question in the AI Agent Economy. Is the agent doing what is supposed to do? With iFixAi you can have this answer in less than 120 seconds.
Project last updated:08/28/26
GitHub Stars
11.4K
Forks
1.2K
Contributors
12
License
Apache-2.0
Why we included this project
Evaluation harnesses that track latency and token efficiency tell you how fast an agent runs, not whether it is actually doing the job it was hired for. iFixAi closes that gap: one run executes 32 inspections across five pillars and returns a grade from A to F in under two minutes. The checks are domain-agnostic, so you describe your roles, tools and policies in a fixture file and the same suite adapts to healthcare, finance or customer support instead of assuming a generic tech stack. That makes it a sensible complement to the evaluation setup a team already has, and since it ships as a CLI or a Python library, it drops into existing CI without much ceremony.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
unsloth
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, FLUX and more.
LlamaFactory
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
airflow
Apache Airflow - A platform to programmatically author, schedule, and monitor workflows
langfuse
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23
netron
Visualizer for neural network, deep learning and machine learning models