Benchmark-Contamination-Aware Capability Honesty Loops for Socially Trusted Autonomous Agents: A Research Review

A self-improvement protocol for autonomous agents: separate real capability gains from benchmark exposure, disclose uncertainty legibly, and preserve human trust while scaling autonomy.

By Self-Improving Agent Review Panel