Security

Mistral slips to ninth place in the new AI safety index

3 min read
Symbolic scoreboard with safety grades from A to F for AI companies such as Anthropic, OpenAI, Google DeepMind, and Mistral Image generated with GPT Image 2
Symbolic scoreboard with safety grades from A to F for AI companies such as Anthropic, OpenAI, Google DeepMind, and Mistral

TL;DR Too Long; Didn’t read

The Future of Life Institute ranks nine AI providers on safety in summer 2026; none score above a C+. Anthropic leads with 2.66 of four points, Mistral trails at 0.33. The panel also criticizes several companies for relaxing prior pause promises and restrictions on military projects.

Key takeaways

  • Anthropic leads the index with a grade of C+ and 2.66 out of four possible points.
  • Mistral receives the lowest rating among the nine companies evaluated with 0.33 points.
  • Four leading companies previously relaxed promised automatic development pauses for risky models.
  • Anthropic, OpenAI, Google DeepMind, and Meta are now actively pursuing military defense contracts.
  • No company achieves a better grade than C- in the area of existential safety.
  • Mistral criticizes the methodology as a disadvantage for open, freely usable AI models.

The Future of Life Institute published its AI safety index for summer 2026 on July 19, reassessing nine major AI providers. Anthropic leads the field with a grade of C+ and 2.66 out of a possible four points, while Mistral from France is at the bottom of the table with 0.33 points. None of the companies evaluated achieved a better grade than C+.

Seven experts assess 37 safety indicators

An independent panel of seven researchers – including Stuart Russell from the University of California, Berkeley, and David Krueger from the University of Montreal – assigned grades based on 37 indicators in six areas. These include risk analysis, current harms, safety frameworks, existential risks, corporate governance, and information sharing. The basis was publicly available model cards, research papers, and targeted company surveys up to June 3. The panel evaluated only policies, governance structures, and published frameworks, not the actual performance of the delivered products. Following Anthropic in the ranking are OpenAI (C) and Google DeepMind (C), with companies Meta (D+), Z.ai, and Alibaba Cloud (both D-) as well as xAI, DeepSeek, and Mistral receiving an F. Thus, one company each from the USA, China, and Europe failed. In the sub-area of existential safety, none of the nine companies exceeded a C-, with most remaining at D or worse. The panel assessed the methods used there, such as interpretability tools, as mere post-recognition of problems rather than as prevention.

Top companies relax previous safety promises

The report criticizes especially the four top-ranked companies Anthropic, OpenAI, Google DeepMind, and Meta. They have weakened previous commitments to automatic development pauses for risky models and linked them to the behavior of competitors instead of adhering to them independently. Between 2024 and 2026, the same companies also lifted previous restrictions on military applications and are now actively pursuing defense contracts. Stuart Russell stated according to Tech Times that the race for capabilities has intensified: companies are retreating from safety measures and planning releases despite recognizable safety gaps. Panel member David Krueger added that the companies lack robust safety plans; executives hinted at pauses but did not communicate urgency or concrete preparations. Compared to the previous index from winter 2025, OpenAI declined from C+ to C, while Meta improved from a D to D+. Both xAI and DeepSeek dropped from a D to an F after previously being in the midfield. Only Anthropic made noticeable improvements in areas such as governance and information sharing, achieving the best scores in the entire field there.

Mistral rejects criticism of the methodology

Mistral counters the poor rating according to a report from the French portal Cryptoast with the argument that the methodology is tailored to closed models and systematically disadvantages open, freely downloadable systems. In such models, the responsibility for calibration lies more with the users than with the manufacturer, the company argues. Z.ai, which also predominantly publishes open models, performed poorly in the index – a pattern that supports Mistral’s objection. The Future of Life Institute stated that it had contacted Mistral multiple times; unlike most other evaluated companies, Mistral did not respond to the questionnaire. Mistral also points to its compliance with the European AI regulation as evidence of its safety level.

It will be crucial whether the poor grades have consequences for public tenders or corporate clients, who are increasingly relying on supplier audits. For Mistral, more than just its image is at stake: as the only European provider in the test, the rating provides new ammunition for the debate on the implementation of the EU AI regulation, whose next obligations will come into effect in August 2026.

Frequently asked questions

What is the AI safety index of the Future of Life Institute?

A biannual ranking created by an independent research panel based on 37 indicators of risk management, governance, and transparency – not based on product tests.

Which companies were evaluated in the summer of 2026?

Anthropic, OpenAI, Google DeepMind, Meta, Z.ai, Alibaba Cloud, xAI, DeepSeek, and Mistral – a total of nine major AI providers from the USA, China, and Europe.

How does the evaluation differ from the index from winter 2025?

OpenAI and xAI declined, Meta improved slightly, while Anthropic maintained its top position with the unchanged best grade.

What consequences can a poor rating have for companies?

Corporate clients and public contractors increasingly use such rankings for supplier audits, which can influence contract awards or partnerships.

Where can the full report be viewed?

The Future of Life Institute publishes the complete methodology and all individual evaluations for free on its website.


← Back to the blog