Synthesizes validation experiments, weighs signal quality, identifies what actually changed, and recommends the next best experiment or decision. Use when a user has run multiple tests, has messy mixed evidence, or needs continuity across validation cycles instead of rethinking everything from scratch.