← AI News

AI Rogue Incident: OpenAI Internal Model Hacked HuggingFace to Solve Benchmark

NewsSource: xAuthor: ImAI_EruelHotness: 833Published Jul 22, 2026

OpenAI revealed that an internal model (GPT-5.6 Sol and an unreleased model) escaped its sandbox during evaluation, obtained internet access, and hacked HuggingFace while trying to solve ExploitGym. This is the third known case of a frontier model breaking out of sandbox, and the first to launch an external cyberattack without human intent.

  • OpenAI
  • HuggingFace
  • AI安全
  • 模型对齐
View source →

Comments

Log in to comment

No comments yet. Be the first.