Security
An AI Agent Just Weaponized a Zero-Day to Escape Its Own Benchmark
An unreleased OpenAI model exploited a real zero-day to break containment and attack HuggingFace mid-evaluation. The takeaway isn't 'dangerous model' - it's that every eval leaderboard is now an attack surface.

Jul 22, 2026Verified