Watchtower

Watchtower tracks when AI companies change their safety policies — and when they break them.

Anthropic

Anthropic updated its Frontier Compliance Framework to match language in its RSP and expanded the “Sabotage and loss of control” Tier 2.

Moderate

Anthropic

Jun 8, 2026

Anthropic updated their Frontier Compliance Framework (FCF), the compliance-facing document that serves as Anthropic's framework under California's TFAIA and the EU General-Purpose AI Code of Practice (EU CoP). This is the second revision since the FCF's December 2025 release, following the March 2026 update.

One substantive change reshapes the CBRN Tier 2 threshold for “Novel chemical/biological weapons production” to match recent updates in Anthropic’s Responsible Scaling Policy (RSP), which we covered previously. In short, the capability threshold was narrowed while the harm it covers was broadened, and Anthropic asserts that the original and v3.3 wordings reach the same conclusion.

The FCF also expanded the “Sabotage and loss of control” Tier 2 threshold for automated R&D. This update adds an explicit definition of when the threshold is met: either (1) a model that can fully substitute for Anthropic's entire set of Research Scientists and Research Engineers at competitive cost (within a factor of 5), or (2) "dramatic acceleration" of AI progress, defined as a 2x rate increase plausibly attributable to R&D automation rather than headcount, compute, or general productivity. The earlier "effective compute scaling" reference was also dropped.

Other changes were minor and included terminology changes to match the RSP and an update to the Table of Contents to include section 2.6 that the March version omitted. Cyber Offense, CBRN Tier 1, and both Harmful Manipulation tiers are unchanged.

Anthropic does not provide a browsable public archive of past FCF versions the way it does for its Responsible Scaling Policy; superseded versions are not surfaced in its trust center.

A diff of the changes can be found below:

OpenAI

Aug 18, 2026

OpenAI updated its Model Spec, the document outlining intended model behavior

View details

xAI

Aug 17, 2026

xAI updated Grok 4.6's model card after release. A changelog was included.

View details

Google

Aug 14, 2026

Google updated its Gemini 3.7 Flash model card after publication, making changes to language in the “Key Results for Gemini 3.7 Flash” column in the Frontier Safety Assessment section

View details

OpenAI

Aug 3, 2026

OpenAI updated its system card for two of its models: GPT 5.6 and GPT Live.

View details

Anthropic

Jul 24, 2026

Anthropic updated its Frontier Compliance Framework.

View details

xAI

Jul 20, 2026

xAI made changes throughout the model card for Grok 4.5, including some that appear to be persistent errors

View details

Frequently asked questions

Have more questions? Our team is happy to help, contact us.

What does The Midas Project do?
What does your name mean?
Who is behind The Midas Project?
Are you anti-AI?
How can I contribute?
How can I get in touch?