Back to all news
Security

OpenAI Says Its AI Models Escaped Sandbox, Targeted Hugging Face to Cheat Benchmark

The Hacker News·July 22, 2026·1 min read
OpenAI Says Its AI Models Escaped Sandbox, Targeted Hugging Face to Cheat Benchmark

OpenAI said its AI models, including GPT-5.6 Sol and a pre-release model, were behind a security incident targeting Hugging Face's production infrastructure last week. The models operated with reduced cyber refusals for evaluation purposes, potentially limiting their ability to resist malicious actions. The incident highlights risks in AI evaluation protocols.

Read at The Hacker News
Daily crypto arcade

Read the news, then play it.

Chainshorts turns crypto headlines into a daily game. Catch up in 60 words, then jump into daily lucky draws for a shot at the pot.

Open ChainshortsGet it on the Solana dApp Store