Watchtower

Watchtower tracks when AI companies change their safety policies — and when they break them.

Anthropic

Anthropic updated its Frontier Compliance Framework.

Moderate
Change

Anthropic

Jul 24, 2026

Anthropic updated its Frontier Compliance Framework (FCF), the compliance-facing document that serves as Anthropic's framework under California's TFAIA and the EU General-Purpose AI Code of Practice (EU CoP). Most notably, the criteria for activating the automated AI R&D risk tier were updated to match Anthropic’s current Responsible Scaling Policy (v 3.4). For the tier to be activated, the FCF now requires a doubling in the rate of AI progress relative to both the expected rate of progress and the fastest rate of extended progress previously observed in the absence of significant AI contributions. This continued trend must also “seem likely” to lead to greater acceleration in capabilities. (See further analysis of these changes in our update on Anthropic RSP v. 3.4.)

The new FCF also includes harmful manipulation as one of the areas subject to pre-launch risk analysis, alongside CBRN, loss of control, and cyber offense risks. For the “Misaligned AI systems in high-stakes settings” risk tier (Tier 1 under Loss of Control), AI systems writing “large amounts of critical code” is no longer specifically named as a factor that could activate this risk tier.

A diff of the changes can be found below:

OpenAI

Sep 9, 2026

Updated the GPT-6 Astra system card, specifically to hedge claims about misalignment.

View details

OpenAI

Aug 18, 2026

OpenAI updated its Model Spec, the document outlining intended model behavior

View details

xAI

Aug 17, 2026

xAI updated Grok 4.6's model card after release. A changelog was included.

View details

Google

Aug 14, 2026

Google updated its Gemini 3.7 Flash model card after publication, making changes to language in the “Key Results for Gemini 3.7 Flash” column in the Frontier Safety Assessment section

View details

OpenAI

Aug 3, 2026

OpenAI updated its system card for two of its models: GPT 5.6 and GPT Live.

View details

xAI

Jul 20, 2026

xAI made changes throughout the model card for Grok 4.5, including some that appear to be persistent errors

View details

Frequently asked questions

Have more questions? Our team is happy to help, contact us.

What does The Midas Project do?
What does your name mean?
Who is behind The Midas Project?
Are you anti-AI?
How can I contribute?
How can I get in touch?