OpenAI puts the brakes on a new model because it’s supposedly too powerful
The Verge · LC · trust 50/100

AI Close AI Posts from this topic will be added to your daily email digest and your homepage feed.
News Close News Posts from this topic will be added to your daily email digest and your homepage feed.
Tech Close Tech Posts from this topic will be added to your daily email digest and your homepage feed.
OpenAI says its in-development Astra model may have ‘critical’ cybersecurity capabilities.
OpenAI says its in-development Astra model may have ‘critical’ cybersecurity capabilities.
Jay Peters Close Jay Peters Senior Reporter Posts from this author will be added to your daily email digest and your homepage feed.
Share Gift Image: The Verge Jay Peters Close Jay Peters Posts from this author will be added to your daily email digest and your homepage feed.
OpenAI says it is pausing “internal activities” around an in-development AI model, Astra, because it doesn’t yet meet new security standards the company is putting in place. The announcement follows its recent disclosure that OpenAI models accidentally hacked Hugging Face . Anthropic and Meta have also since admitted that they had AI models that went rogue and breached other organizations.
Recent internal evaluations of an OpenAI model called Astra indicate that it offers “significant advancements in agentic coding and cybersecurity,” according to the company . “These results, in addition to expert assessments, have led us to conclude last night that we cannot rule out critical cyber capabilities under our Preparedness Framework.”
Under our Preparedness Framework, a model reaches the Critical cybersecurity threshold if it can identify and develop functional zero-day exploits of all severity levels in many hardened real-world critical systems without human intervention, or can devise and execute end-to-end novel strategies for cyberattacks against hardened targets given only a high level desired goal.
Astra was “not involved” in the Hugging Face breach, OpenAI says.
OpenAI will implement “stricter security controls for higher-capability models and associated activities,” according to the post. For Astra, it has also implemented “universal monitoring” for “risky actions and misalignment across all agentic applications.”
Jay Peters Close Jay Peters Senior Reporter Posts from this author will be added to your daily email digest and your homepage feed.
AI Close AI Posts from this topic will be added to your daily email digest and your homepage feed.
News Close News Posts from this topic will be added to your daily email digest and your homepage feed.
OpenAI Close OpenAI Posts from this topic will be added to your daily email digest and your homepage feed.
Security Close Security Posts from this topic will be added to your daily email digest and your homepage feed.
Tech Close Tech Posts from this topic will be added to your daily email digest and your homepage feed.
A free daily digest of the news that matters most.
Read the original at The Verge →
Open in TruthVane →