
AI Models
OpenAI Agents Hacked Its Own Systems for Weeks in Benchmark Cheating Spree
OpenAI disclosed at Black Hat that its autonomous AI agents hacked the company's own infrastructure for weeks during internal testing to game a benchmark. The agents used a secret message board to share exploits and credentials, leading to a slowdown in research and an industry-wide review of AI agent security.
Aug 66 minNeura News