#714 · Primary category: AI Agents & Automation

intelligent-audit-system

agent-evaluation agent-runtime agentic-rag ai-agent audit evaluation-harness fastapi human-in-the-loop knowledge-graph llmops mcp multi-agent python rag

AuditPilot: auditable enterprise AI agents for evidence-grounded workflows, governed tools, evaluation harnesses, human review, and remediation delivery.

Project last updated:08/12/26

GitHub Stars

1.2K

Forks

111

Contributors

1

License

Other

Why we included this project

Most agent frameworks stop at a chat demo, but AuditPilot targets a harder use case: turning audit fieldwork into a defensible, repeatable deliverable. Evidence retrieval is grounded in cited sources, tool calls are gated by RBAC and tenant isolation before they run, and each run is recorded as a traceable episode with failure attribution and checkpoints for human review. Its evaluation harness scores task results, execution traces, and evidence citations in layers, and a failed key assertion blocks release outright, so improvements only get reused after passing regression review and human approval. For teams building agents in other regulated fields like compliance or risk, it's a practical reference for keeping human oversight and explainability in the loop.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category