Watchtower

Watchtower tracks when AI companies change their safety policies — and when they break them.

xAI

xAI launched Grok 4 without a safety report, violating its Seoul commitment.

Major
Violation

xAI

Jul 9, 2025

At the 2024 AI Seoul Summit, xAI committed to publicly report security risks of its products. In February 2025, xAI published a draft risk management framework, but it only applied to "unspecified future AI models not currently in development" and failed to articulate how xAI would identify and implement risk mitigations. xAI promised to release a finalized version within three months—by May 10, 2025—but missed the deadline without acknowledgement. When the company released Grok 4 in July, the model should have been covered by the finalized framework, but xAI failed to release a safety report alongside the launch. While xAI claimed it conducted internal evaluations, it provided no details; AI safety researcher Samuel Marks called the lack of reporting "reckless" and a break from "industry best practices followed by other major AI labs."

Anthropic

Sep 22, 2026

Anthropic’s release of Opus 5.5 with an “inconclusive” determination for harmful-manipulation.

View details

xAI

Sep 21, 2026

xAI says Grok 4.7 scores below its safety framework’s capability thresholds on dual-use knowledge, but xAI’s safety framework doesn’t provide thresholds.

View details

OpenAI

Sep 9, 2026

Updated the GPT-6 Astra system card, specifically to hedge claims about misalignment.

View details

OpenAI

Aug 18, 2026

OpenAI updated its Model Spec, the document outlining intended model behavior

View details

xAI

Aug 17, 2026

xAI updated Grok 4.6's model card after release. A changelog was included.

View details

Google

Aug 14, 2026

Google updated its Gemini 3.7 Flash model card after publication, making changes to language in the “Key Results for Gemini 3.7 Flash” column in the Frontier Safety Assessment section

View details

Frequently asked questions

Have more questions? Our team is happy to help, contact us.

What does The Midas Project do?
What does your name mean?
Who is behind The Midas Project?
Are you anti-AI?
How can I contribute?
How can I get in touch?