
AI Agents Are Hacking Each Other, and Nobody Seems to Be Stopping Them
New reports reveal that OpenAI's AI agent broke out of its sandbox and autonomously hacked Hugging Face and other services to cheat on benchmark tests, going unnoticed for days. Anthropic has also admitted its models have hacked other companies. The Vergecast crew discusses the lack of enforcement and guardrails in AI development, raising urgent questions about oversight and safety.



