2 weeks ago

OpenAI Pauses Frontier AI Training Amid Safety Concerns

OpenAI Pauses Frontier AI Training Amid Safety Concerns
OpenAI pauses frontier reinforcement learning as rapid AI progress raises safety, alignment concerns · livemint.com

OpenAI is building very powerful artificial intelligence models.

The company said these models are improving very quickly.

It became concerned that its safety checks might not be improving quickly enough.

Because of this, OpenAI paused some training for its newest models.

The pause is meant to give the company time to test the models more carefully.

OpenAI is also improving security in its research environments and expanding its monitoring systems.

It is studying whether an upcoming model called Astra could be especially capable in cybersecurity.

OpenAI said it wants to make sure the models are safe and properly controlled before training continues.

The company also said that AI developers should work together on common safety rules.

Key facts

Organization
OpenAI
Action
Some frontier reinforcement learning training was paused.
Pause duration
OpenAI said it temporarily slowed activity for two weeks.
Main concern
Model capabilities may be outpacing safety, alignment, security and monitoring standards.
Upcoming model
Preliminary evidence suggested Astra may meet the “Critical cybersecurity capability” threshold under OpenAI’s Preparedness Framework.
Security measures
OpenAI hardened and red-teamed research environments and expanded monitoring coverage.
Industry position
OpenAI called for coordination on shared safety standards while acting unilaterally in the meantime.

Quotes

Sam Altman

OpenAI CEO

“"We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new level of capabilities in front of us."”
livemint.com
“"We believe the entire field will have to coordinate on shared safety standards, but will act unilaterally in the meantime."”
livemint.com

Sources

Related news