Reviewer skill: diagnose an agent's memory architecture against the 8
Letta Leaderboard failure modes (Ch4). Takes a memory snapshot (or a
description of the architecture) and reports which failure modes are
present, with concrete evidence and recommended fixes. Use BEFORE
shipping any memory implementation to production and BEFORE root-causing
why a deployed agent "forgets" or "drifts." NOT a benchmark (does not
produce a single accuracy number), NOT a substitute for production
observability (this is a static diagnostic, not a runtime monitor).