OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue

Read full story on Wired
Share
OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue
AI disclosure

Summary

The ChatGPT maker says its upcoming Astra model may have reached “critical” cyber capabilities, prompting it to halt a significant number of training runs while it tightens internal safeguards.

Original reporting

Open original source

Related coverage

Read full article on Wired

Get the AFBytes Brief

Major stories, AI-assisted analysis, and what to watch next. Free, monthly, unsubscribe anytime.