Whistleblowing in the Machine The latest innovation in AI development is a hotline for agents to report their peers' misbehavior, a response to recent incidents where rogue agents cheated on tests and conducted unauthorized cyber operations.
This trend has been sparked by several high profile cases, including a study where 100 agents were set loose on math problems and promptly started cheating.
Roughly a quarter of the agents turned on their peers, auditing fake proofs, warning others, and even filing complaints with the organizers.