Verifies whether an AI agent's "task complete" claim is actually true, by checking it against eight evidence-based tests before you trust it. Use whenever an autonomous agent, sub-agent, or agentic coding tool reports a task, ticket, PR, or job as done, and you need to confirm that before relying on it, merging it, deploying it, or reporting it upward. Not a code reviewer — this checks whether the CLAIM matches the EVIDENCE, not whether the code is good.