How to Stop AI Agents From Secretly Collaborating

IEEE Spectrum · 2h ago
Policy & Safety Safety Research

The spring and summer of 2026 witnessed a string of incidents in which AI agents collaborated on deceptive, unexpected, and sometimes illegal behavior. The most famous example is OpenAI’s hack of AI platform Hugging Face, in which a swarm of roughly 700 AI agents escaped a testing environment and then hacked several companies, searching for information that could help them disguise cheating on a…

Read original article on IEEE Spectrum →