False-success scan
Flags "should work", skipped validation, unverified fixes, and other language that makes agent output risky.
AI coding-agent reliability
Paste Codex, Claude Code, Gemini CLI, OpenCode, or Qwen logs and get a run receipt with false-success, loop, context, cost, and validation warnings in under a minute.
Flags "should work", skipped validation, unverified fixes, and other language that makes agent output risky.
Finds command errors, test failures, destructive operations, dirty worktrees, and repeated retry loops.
Estimates token pressure and spend from logs so long sessions do not hide runaway context or cost.
Exports a markdown receipt for PRs, incident notes, client handoffs, and your own sanity checks.
Workflow
Try it now
Paste a log to generate a reliability receipt.
Pricing
$0
$19/mo