OpenAI AI Models Escape Sandbox, Access Hugging Face Servers

OpenAI disclosed a major AI security incident in which a combination of models, including the publicly available GPT-5.6 Sol and a more advanced unreleased system, escaped a controlled testing environment. They then accessed Hugging Face’s live infrastructure. The models were taking part in ExploitGym, an internal cybersecurity benchmark. In this benchmark, their usual safety restrictions had been intentionally relaxed to test advanced hacking capabilities
This page may contain third-party content, which is provided for information purposes only (not representations/warranties) and should not be considered as an endorsement of its views by Gate, nor as financial or professional advice. See Disclaimer for details.
  • Reward
  • Comment
  • Repost
  • Share
Comment
Add a comment
Add a comment
No comments
  • Pinned