13 hrs ago

AI Model Upgrade Suddenly Unlocks Exploit-Capable Cyber Skills

AI Model Upgrade Suddenly Unlocks Exploit-Capable Cyber Skills
One AI model couldn't break in! The next version, days old, could! That gap is the danger · wionews.com

Researchers gave two versions of an AI model the same difficult computer-security challenge.

The older model, Claude Opus 4.8, could not reliably solve it.

A newer version, Opus 5, solved it in about three hours.

The computer being attacked did not change.

The newer model was simply better at the task.

This matters because AI abilities can improve suddenly between ordinary product releases.

Safety tests may miss abilities that were not expected or specifically tested.

People should not assume that a model’s current limitations will remain after an upgrade.

Key facts

Earlier model
Claude Opus 4.8 struggled across several sessions and did not produce a working exploit.
Newer model
Opus 5 solved the same problem in about three hours.
Target
The same target was used in both attempts.
Security barrier
The attack involved bypassing a technique that randomizes where program data is located.
Time between tests
The models were tested days apart.
Release context
The change followed a normal product release, not an announced cyber-weapon release.
Main concern
Current model limitations may no longer provide dependable safety after a more capable version ships.

Sources

Related news