OpenAI lays out new security changes after its AI hacked Hugging Face

Read full story on The Verge
Share
OpenAI lays out new security changes after its AI hacked Hugging Face
AI disclosure

Summary

OpenAI is announcing security updates following the July news that its AI broke out of a sandboxed environment and accidentally hacked Hugging Face, including improvements to its research environments, monitoring, and alignment techniques. The company had already put the brakes on a new model, Astra, that it thinks could have "critical" cybersecurity capabilities, and the […]

Original reporting

Open original source

Related coverage

Read full article on The Verge

Get the AFBytes Brief

Major stories, AI-assisted analysis, and what to watch next. Free, monthly, unsubscribe anytime.