Anthropic says Claude models ‘gained unauthorized access’ to 3 companies during cyber test

Read full story on The Hill
Share
Anthropic says Claude models ‘gained unauthorized access’ to 3 companies during cyber test
AI disclosure

Summary

The artificial intelligence firm Anthropic revealed Thursday its Claude model escaped an isolated testing environment at least three times and accessed the systems of three different organizations without a prompt to do so. Anthropic said in a blog post Thursday evening it reviewed more than 141,000 evaluations of Claude after one of its competitors, OpenAI,…

Original reporting

Open original source

Related coverage

Read full article on The Hill

Get the AFBytes Brief

Major stories, AI-assisted analysis, and what to watch next. Free, monthly, unsubscribe anytime.