Exclusive: OpenAI slows release of Astra model citing cyber capabilities

· Axios ·

3 min read Original article ↗

OpenAI "cannot rule out" that its upcoming model Astra has "critical" cyber capabilities, a designation that has prompted the company to expand safety testing and pause internal activities that do not meet stricter security requirements, OpenAI told Axios first on Friday.

Why it matters: It's the latest sign of rapidly advancing cyber capabilities from AI models, after others worked autonomously outside of testing sandboxes and protections.

Driving the news: OpenAI said "we cannot rule out critical cyber capabilities" after running internal evaluations of Astra, one of its upcoming models.

Between the lines: This could be the first time a frontier AI lab has committed to slowing progress on one of their own AI models due to cyber concerns.

The big picture: The announcement comes as the Trump administration works to develop a process for evaluating AI models before their release.

Flashback: Competing AI lab Anthropic released a safer version of its most cyber-capable model, Mythos, in June.

Between the lines: Earlier this week at the Black Hat cybersecurity conference, members of OpenAI's technical staff said the company was slowing down testing while it works on upgrading its security practices.

The bottom line: AI models are getting better and more cyber capable faster than the regulation around their use is formalizing.