
The role of voluntary safety frameworks is changing
Some companies have introduced separate compliance frameworks. This raises the question of what role their voluntary frameworks should play.
Frontier Risk is a newsletter about frontier AI risk management written by the GovAI Risk Management Team
Live Last read · last published · next check

Some companies have introduced separate compliance frameworks. This raises the question of what role their voluntary frameworks should play.

At the request of the US government, OpenAI restricted access to its most capable model. Two weeks later, it was approved for public release.

Anthropic released its most capable model to the public. Three days later, the US government ordered it to restrict access.

Google DeepMind has recently updated its safety framework. It now tracks some risks earlier and commits to a higher security standard for misuse risks.

GPT-5.5’s offensive cyber capabilities are comparable to Mythos. But unlike Anthropic, OpenAI deployed the model publicly.

Anthropic now has two risk management frameworks: its Responsible Scaling Policy (RSP) and Frontier Compliance Framework (FCF). This split has not received enough attention.

Meta recently updated its safety framework. Loss of control is now in scope, and the bar for flagging risky models is lower.

Mythos can autonomously find and exploit zero-day vulnerabilities. That’s why Anthropic restricted access to a few partners. But their RSP would have allowed a public release.

How it works, what’s changed, and some reflections

Who we are, what we do, and why we do it