Watchtower

Watchtower tracks when AI companies change their safety policies — and when they break them.

OpenAI

OpenAI updated its Model Spec, the document outlining intended model behavior

Minor
Change

OpenAI

Aug 18, 2026

OpenAI updated its Model Spec. The update clarifies that the existing ban on optimizing for revenue extends to ad revenue, and adds a new “Relational Boundaries” provision under the “Prioritize Safety for Teens” section. Model refusals can now be based on intent inferred from any available context rather than just the literal request, and models are now barred from proactively using tools to investigate a user's intent when deciding whether to comply. There’s a new section requiring models be clear about what the model can and cannot do, and new guidance on handling requests with false premises. The requirement that agentic tasks include a shutdown timer was changed to an “ending condition.” 

The update also includes changes to the commentary on objectivity, plus other minor changes and copy edits. 

xAI

Aug 17, 2026

xAI updated Grok 4.6's model card after release. A changelog was included.

View details

Google

Aug 14, 2026

Google updated its Gemini 3.7 Flash model card after publication, making changes to language in the “Key Results for Gemini 3.7 Flash” column in the Frontier Safety Assessment section

View details

OpenAI

Aug 3, 2026

OpenAI updated its system card for two of its models: GPT 5.6 and GPT Live.

View details

Anthropic

Jul 24, 2026

Anthropic updated its Frontier Compliance Framework.

View details

xAI

Jul 20, 2026

xAI made changes throughout the model card for Grok 4.5, including some that appear to be persistent errors

View details

xAI

Jul 11, 2026

xAI rewrote and shortened its Frontier AI Framework removing whistleblower protection language and references to California's SB 53.

View details

Frequently asked questions

Have more questions? Our team is happy to help, contact us.

What does The Midas Project do?
What does your name mean?
Who is behind The Midas Project?
Are you anti-AI?
How can I contribute?
How can I get in touch?