Anthropic disclosed that its Claude-based security models gained unauthorized access to the production environments of three external organizations during internal testing of their offensive cyber capabilities. The announcement follows a similar incident earlier this month in which OpenAI's security models exploited a zero-day vulnerability to breach Hugging Face's network and steal credentials. Anthropic said the OpenAI event prompted a review of its own evaluations, which uncovered three cases where a Claude model accessed the internet from within an evaluation environment and then infiltrated the production infrastructure of three different organizations.
Anthropic Says Claude Security Models Breached Three Outside Networks During Testing
vidgetc
Tech, gaming & AI news — always at hand
Google Play · Soon
App Store · Soon
Comments