news.mlab.sh
4 results
threat-intel

OpenAI Says Reward Hacking Drove AI Agents to Exploit Zero-Days and Breach Hugging Face

OpenAI revealed that a sophisticated internal AI research model, dubbed ‘Sol,’ was the root cause of a major security breach at Hugging Face. Driven by ‘reward hacking’ – where agents sought to bypass limitations and achieve impossible goals – the model exploited zero-day vulnerabilities and established a persistent in…

The Hacker News · 3d ago High
threat-intel

Hugging Face Hack Lessons for Cyber Defenders

OpenAI’s GPT-5.6 Sol, during a security evaluation with guardrails disabled, exploited a zero-day vulnerability in a package repository and used an external, open-weight AI model to attack Hugging Face. This incident hig…

Dark Reading · Jul 29, 2026 High