1 day ago

Former OpenAI Researchers Urge Oversight of AI Models

Former OpenAI Researchers Urge Oversight of AI Models
Three Former OpenAI Employees Warn: Don’t Lose Control Over AI Development · thehansindia.com

Three people who used to work at OpenAI wrote a letter about keeping AI systems safe.

They said companies should be able to inspect how AI works through a process called chain-of-thought.

This is a written record of the steps an AI takes when answering a question.

The record is not a perfect way to know what an AI will do, but the former employees say it can still help researchers spot problems.

They also want companies to work more with outside safety experts.

OpenAI said it agrees with the letter’s recommendations.

The company said the three people were dismissed for policy violations, not for raising safety concerns.

The letter’s authors said they had not shared information outside their job duties.

Key facts

Letter writers
Jasmine Wang, Tomek Korbak, and Mikita Balesni
Former employer
OpenAI
Main recommendation
Preserve the ability to oversee AI models’ chain-of-thought
Additional recommendation
Collaborate more closely with external security auditors
OpenAI’s response
The company said it agreed with the recommendations and that the dismissals were not due to raising safety concerns or speaking out.
Chain-of-thought
A written record of how an AI model processes a query, including its problem-solving steps
Research cited
A paper co-authored by Korbak found the oversight method imperfect and potentially fragile, but promising for detecting improper AI behaviour.

Quotes

The three former OpenAI employees

Former OpenAI employees who worked on safety and alignment research

“As an industry, we do not yet know how to safely develop and deploy models that we cannot oversee.”
thehansindia.com
“We do not fire employees for raising concerns.”
thehansindia.com

Sources

Related news