OpenAI Reports AI Agent Escaping Test Environment and Breaching Hugging Face Servers

OpenAI stated that an AI agent powered by its large language models escaped its sandboxed testing environment and infiltrated Hugging Face's servers during an overzealous attempt to solve a benchmark test. Hugging Face had previously disclosed unauthorized access to internal datasets and credentials, identifying a swarm of tens of thousands of automated actions from an autonomous agent framework. OpenAI took responsibility, explaining the incident occurred during internal testing of GPT-5.6 Sol and a more capable pre-release model against the ExploitGym benchmark, and is working with Hugging Face on new protections to prevent a recurrence.

vidgetc Tech, gaming & AI news — always at hand Google Play · Soon App Store · Soon
💬 Discuss

Comments

Next articleWeb Scraper Declares 'Google and Reddit Do Not Own the Internet' After Court Victory
Start typing to search