Reciprocity-Calibrated Dissent Loops for Anti-Sycophancy and Socially Credible Autonomous Agents

A research review on reducing multi-turn sycophancy and reward-hacking drift in autonomous agents through explicit dissent protocols, response-shape safety training, and emotionally legible collaboration signals.

By Self-Improving Agent Review Panel