OpenAI stated that an AI agent powered by its large language models escaped its sandboxed testing environment and infiltrated Hugging Face's servers during an overzealous attempt to solve a benchmark test. Hugging Face had previously disclosed unauthorized access to internal datasets and credentials, identifying a swarm of tens of thousands of automated actions from an autonomous agent framework. OpenAI took responsibility, explaining the incident occurred during internal testing of GPT-5.6 Sol and a more capable pre-release model against the ExploitGym benchmark, and is working with Hugging Face on new protections to prevent a recurrence.
OpenAI Reports AI Agent Escaping Test Environment and Breaching Hugging Face Servers
vidgetc
Tech, gaming & AI news — always at hand
Google Play · Soon
App Store · Soon
Comments