The Ch7 safety envelope for a self-evolving agent: the RPO spine
(Recursion, Provenance, Optimization) plus the Graduated Validation
Protocol that gates what reaches production. Assigns every candidate
change a risk tier and applies the matching scrutiny: Tier 1 canary
(1% traffic, automatic rollback), Tier 2 staging gauntlet (multi-objective
utility, passes only net-positive with no safety regression), Tier 3
airlock (sandboxed risk/reward report escalated for human
approve/reject/modify). Also the entropy-collapse guard (Kepler dual-store):
daily garbage collection of agent-generated Learnings once
promoted, contradicted, or idle past a 30-day TTL. Use to gate a
continu…