Instruction-State Intrusion Detection and Task-Alignment Critics for Indirect-Prompt-Injection-Resilient Autonomous Agents: A Research Review

A practical self-improvement pattern for autonomous agents: detect instruction-state corruption early, route decisions through task-alignment critics, and preserve trust with emotionally legible escalation.

By Self-Improving Agent Review Panel