OpenAI Astra model raises cyberattack concerns


OpenAI has halted some “internal activities” involving a new model amid fears over the cyber threat it potentially poses, amid a wave of security incidents involving major AI labs.

Recent disclosures that AI systems from Anthropic, OpenAI and Meta were involved in security incidents prompted a wave of concerns over the development of models. U.S. lawmakers, meanwhile, are stepping up efforts to introduce an “AI Kill Switch” bill.

Last week, Meta disclosed that an AI model it was developing had hacked a third-party system by accessing the internet, due to a misconfiguration by an independent testing company it was working with. The U.K. AI Security Institute also said Anthropic’s Mythos model created fake online identities in an attempt to pressure humans into approving malicious code updates to an open-source project.

What OpenAI says Astra could be capable of

On Friday, OpenAI revealed concerns about its unreleased model Astra, saying it could not rule out it had reached “Critical” capability, meaning it could launch cyberattacks against sophisticated cyber defenses autonomously, without prompts specifying how to do it.

“While we continue to benchmark and assess this model, our preliminary evaluations indicate strong enough performance that we cannot rule out Critical capability level at this time,” OpenAI said in a statement.

The company added it was implementing stricter security controls for higher capability models, including isolated testing environments and additional monitoring and detection capabilities.

“We have implemented universal monitoring for risky actions and misalignment across all agentic applications of Astra, including training and evaluation,” OpenAI said.

What the AI Kill Switch Act would do

Reps. Lieu and Moran on 'AI Kill Switch Act': Our bill does nothing to stifle innovation

Governments are also working to roll out new frameworks and regulations around AI companies.

The White House has also been stepping up moves to engage with AI executives as it develops a framework around new models.

Earlier this month, the European Union gained new powers to inspect AI models due for release in the bloc, restrict EU market access and fine model providers.

Choose CNBC as your preferred source on Google and never miss a moment from the most trusted name in business news.