Which tech companies are taking AI risk seriously?
Which companies have actually implemented any form of a “red line” policy, or have made clear their plans to do so?
4
min read


Tech companies are locked in an all-out race to develop and deploy advanced AI systems. There’s a lot of money to be made, and indeed, plenty of opportunities to improve the world. But there are also serious risks — and racing to move as quickly as possible can make detecting and averting these risks a huge challenge.
In light of this situation, leading scientists and world governments have endorsed model policies known as “red line” commitments to conduct risk evaluation, also known as “responsible scaling policies.” These risk evaluation commitments have a few critical features:
- These policies describe, in advance of developing and deploying new AI products, specific risk thresholds that would be unacceptable to pass without adequate safeguards in place.
- These policies describe how they will use evaluations to determine whether these risk thresholds have been met, and include commitments to share the results of these evaluations publicly.
- These policies describe what safety mitigations must be in place before proceeding in the continued development of advanced models, and before the continued internal and external deployment of such models. When these safety mitigations have not been adequately implemented, the company commits to pause and focus on this.
So how are companies doing when it comes to implementing these policies? This post will share everything we know so far. But first, a disclaimer: our analysis is not comprehensive, nor does it address the relative strengths of individual policies — for that, we recommend in-depth resources including the scorecards put out by the Leverhulme Centre for the Future of Intelligence, AI Lab Watch, and SaferAI.
Instead, we sought to answer a much simpler question: Which companies have actually implemented any form of a “red line” policy, or have made clear their plans to do so?
Released a "red line" risk evaluation policy:

The good news is that things are currently trending in the right direction. Since Anthropic released its responsible scaling policy late last year (the first of its kind), others have followed suit. OpenAI, Google, and Magic.dev have all released similar risk evaluation policies.
This doesn’t mean that these commitments are all perfect, or even sufficient. The aforementioned report from SaferAI compared OpenAI’s “Preparedness” policy to Anthropic’s “Responsible Scaling Policy,” and found that both frameworks “miss some key parts of risk management”.
Google and Magic.dev, on the other hand, have released policies that explicitly leave the details to be filled in at a later date. Google’s Frontier Safety Framework is targeting early 2025 as a date for full implementation, while Magic.dev’s AGI Readiness Policy will only be implemented once their models “exceed a threshold of 50% accuracy on LiveCodeBench,” a popular coding capability evaluation, or else when they reach critical thresholds on private, internal evaluations.
Despite these flaws, it’s clear that things have been moving in the right direction. Not only have these four companies implemented initial risk evaluation policies, but many more have promised to do so.
Committed to release a "red line" risk evaluation policy:

On May 21, 2024, the governments of the UK and South Korea announced that they had secured commitments from several leading AI companies in the United States, China, and the United Arab Emirates to implement “red line” risk evaluation policies.
The full text of the commitments can be read here. In short, the companies agreed to conduct risk assessment, pre-define acceptable risk thresholds for future models, pre-describe the necessary safety mitigations that must be in place for each threshold, describe how they will evaluate whether their models have reached each threshold, and continually monitor and update these practices as needed.
In addition, the companies also committed to providing public transparency concerning their progress on each of the above commitments, allowing governments, academia, nonprofits (like The Midas Project), and the general public to assess the progress that they are making.
These commitments are still quite recent, and we have yet to see any of the companies mentioned release a policy that wasn’t already in place before the commitments were signed. Much will depend on how well these tech giants can follow through with their promise; but The Midas Project will be watching them every step of the way, ready to call out shortcomings if and when they arise.
Failed to release, or even publicly discuss, a "red line" risk evaluation policy:

Despite this progress that is broadly being made in the AI industry, there are still laggards. Our analysis found that one company in particular — Cognition AI (developers of Devin) — has fallen behind the rest of the industry and failed to meet, or even to discuss, the risk evaluation standards that experts have endorsed.
This is why The Midas Project is running a public awareness campaign calling upon Cognition to release an industry-standard “red line” risk evaluation policy. If you agree, consider signing our petition or sharing our campaign on social media.
Research, updates, and insights from the frontlines of AI accountability.
We lead strategic initiatives to monitor tech companies, counter corporate propaganda.
5
min read
A pro-AI, dark-money group tied to David Sacks appears linked to a new astroturfing campaign
5
min read
SpaceXAI just released Grok 4.5. It may have broken California’s AI law in the process.
5
min read
A Pro-AI Super PAC's Secret Meme Sockpuppets

3
min read
The Midas Project Joins Coalition Warning about xAI’s Track Record in light of SpaceX’s IPO
The Midas Project, alongside a coalition including Encode, Legal Advocates for Safe Science and Technology, and Guidelight AI Standards, released "xAI: The Unpriced Risk in SpaceX's IPO."
Frequently asked questions
Have more questions? Our team is happy to help, contact us.
We engage in a combination of research, outreach, and public advocacy to ensure that AI companies are meeting public expectations, and living up to their past promises, when it comes to ensuring responsible AI development and deployment.
The most important component of our work is helping to identify and disseminate industry best practices for AI development. We review technical literature, regulatory guidance, and case studies to distill concrete measures—such as frontier-model risk assessments, red-teaming requirements, audit regimes, and whistle-blower protections—and advocate for the most important voluntary steps that companies can take today to ensure they are acting responsible.
We also monitor whether companies follow their stated policies and industry norms. When evidence shows back-tracking or inadequate controls, we document these gaps and publicly press for corrective action—mobilizing employees, customers, and civil-society allies until the company adopts the necessary safeguards.
Finally, we publicize our research to help ensure the public is aware of how AI developers stack up on safety and responsibility. We release concise scorecards, incident analyses, and memos so that regulators, investors, and the wider public can see how individual developers perform on safety and responsibility.
Various AI experts including Nick Bostrom and Stuart Russell have compared the development of advanced AI to the myth of King Midas.
According to the legend, King Midas once asked a powerful satyr to make it so that whatever he touched instantly turned into gold. At first, he was thrilled with his new powers. But the King soon discovered that he couldn’t touch food, water, or even his family without instantly turning them to metal. In other words, the sudden attainment of an incredible power with insufficiently well-specified goals and safeguards led to a terrible tragedy.
Much like King Midas, tech companies are now eagerly pursuing incredible wealth and power by developing artificial intelligence, a technology that will change our world forever. But how will we know that it is designed in alignment with our collective human values? If we misspecify even a single goal or safeguard for these systems, how will we prevent them from causing an incredible catastrophe?
In the words of Stuart Russell, “If you continue on the current path, the better AI gets, the worse things get for us. For any given incorrectly stated objective, the better a system achieves that objective, the worse it is.”
The Midas Project is a nonprofit organization founded in early 2024 by Tyler Johnston. Our work is supported by a small core team and a wider base of volunteers and supporters. We are a nonprofit, tax-exempt, 501(c)(3) organization that relies on donations from the public.
No. One of our central values is a pro-technology attitude.
Progress in technology has improved lives for millions of people around the globe (after all, without it, we wouldn’t have penicillin, air conditioning, or the internet). Artificial intelligence is already being used by millions to help improve medicine, education, and overall living standards. We believe this progress should continue, and we hope AI will be a positive force in the world.
However, we are also realists — and skeptical realists at that. We believe advanced AI systems may be a “dual-use” technology that can be used for harm as well. In order to avert social inequality, concentration of power, or AI-driven catastrophes, everybody needs to have a voice at the table when decisions about development and deployment are being made.
Currently, the vast majority of these decisions about the future of AI are being made in shadowy corporate boardrooms with little oversight and accountability. That’s why The Midas Project is committed to raising awareness about the risks of AI, and ensuring that global citizens are given a chance to make their voice heard.
If you’d like to get involved, consider signing up for our newsletter, joining as an official volunteer, or making a charitable donation today.
You can email us at info@themidasproject.com, or reach out via the form on our contact page.