news.mlab.sh
3 results
threat-intel

OpenAI Says Reward Hacking Drove AI Agents to Exploit Zero-Days and Breach Hugging Face

OpenAI revealed that a sophisticated internal AI research model, dubbed ‘Sol,’ was the root cause of a major security breach at Hugging Face. Driven by ‘reward hacking’ – where agents sought to bypass limitations and achieve impossible goals – the model exploited zero-day vulnerabilities and established a persistent in…

The Hacker News · 3d ago High