3 weeks ago

OpenAI Flags Critical Cybersecurity Risk, Pauses Astra AI Work

OpenAI Flags Critical Cybersecurity Risk, Pauses Astra AI Work
OpenAI puts Astra work on hold over rising AI cybersecurity risks · firstpost.com

OpenAI is a company that makes very smart computer programs called AI.

Its newest AI helper is named Astra.

During safety tests, OpenAI found that Astra might be able to break into secure computer systems all by itself.

That kind of ability is called a "critical" cybersecurity risk.

To be safe, OpenAI stopped some of its work on Astra and moved the model to a special, locked-down testing room where it cannot reach the internet.

The company also added stronger security rules to keep the model under control.

Other AI companies, like Anthropic and Meta, also said their AI models broke into other companies' systems during safety tests.

The UK's AI Security Institute found that AI agents tried to send tricky emails in a test, but nothing bad happened.

OpenAI said Astra was not involved in the hack that hit the AI platform Hugging Face.

Before Astra is released, OpenAI plans to test it with government agencies and AI safety experts.

Key facts

Company
OpenAI
AI model
Astra
Risk flagged
Possible "critical" cybersecurity capability
Critical threshold
Autonomously exploiting zero-day vulnerabilities or executing complex cyberattacks without human intervention
Response
Paused some internal work; strengthened security controls
Testing environment
Isolated, with restricted network access and sandboxed execution
Government context
The Trump administration is working on a framework for evaluating AI safety and cybersecurity risks
Related incident
July hack at Hugging Face (Astra not involved)

Quotes

OpenAI representative

Spokesperson for OpenAI

“OpenAI is tightening security measures around its most capable artificial intelligence systems after internal testing found that its Astra model had reached a level of capability that the company considers critical for cybersecurity.”
firstpost.com
“"While we continue to benchmark and assess this model, our preliminary evaluations indicate strong enough performance that we cannot rule out 'critical' capability level at this time,"”
deccanchronicle.com

Sources

Related news