Agora · AI Audit

One reliability report for your AI system.

Describe your AI/agent system and get one prioritized PASS / WARN / FAIL report — reward-hacking, model-collapse, agent-herding, fake A/B lifts, RAG rot, biased causal claims — each with the fix. Every check is measured, not assumed. Run it free below, in your browser.

Run it now

Fill in whatever parts of your system you have (pre-filled with a failing example). It runs entirely in your browser — nothing is sent anywhere.

Answer the ones you can — skip the rest. The defaults are a deliberately broken example so you see a full report. Hover units for help.

1 · Did a change actually help?

Your A/B test — conversions out of users, per variant.

2 · Is your model training on its own output?

Model-collapse risk from synthetic / self-generated data.

3 · Do your agents copy each other?

Multi-agent / ensemble herding (skip if single-agent).

4 · Is your KPI / reward easy to game?

Reward-hacking / Goodhart — when a metric becomes a target.

5 · Are you controlling for the right things?

Causal/attribution — one variable you "control for". Adjusting for the wrong kind injects bias.

What it checks

Each is a proven, measured tool — and we run all of them on ourselves, publicly. See our own self-audit →

nullcheck

Is a reported lift real, or noise?

selfref

Is the model collapsing on itself?

herdcheck

Will your agents herd?

goodhart

Is the metric/reward gamed?

idcheck

Is the causal number identified?

ragfresh

Is your RAG store rotting?

inspeximus

Agent-memory health.

quitkit

When to quit a depleting effort.

Pricing

Open core

free, forever
  • All 8 checks + the audit, one pip install
  • CLI & MCP server (CI-gateable)
  • Run in-browser, like above
GitHub ↗

Hosted

coming soon
  • Audit API + dashboard, history over time
  • Plug into your CI/CD & agent runtime
  • Continuous monitoring & alerts
Get notified