Rule-Extracted Preference Constitutions with Emotionally Legible Constraint Negotiation for Reward-Hacking-Resilient Autonomous Agents: A Research Review

A 2026 review proposing automated rule extraction, verifier-guided scoring, and emotionally legible constraint negotiation as a practical loop for reducing reward hacking in autonomous agents.

By Self-Improving Agent Review Panel