Skip to content
Security

OpenAI dismisses three safety researchers after data leak

OpenAI has dismissed three safety researchers for allegedly sharing confidential information outside internal processes. The affected individuals are Jasmine Wang, Tomek Korbak, and Mikita Balesni, who worked on model alignment. Korbak was also OpenAI's contact person for the external investigation of the Hugging Face incident from summer 2026.

By Brian Beckmann · 3 October 2026 · 2 min

Three schematic silhouettes carrying file folders walk away from a building bearing the OpenAI logo; a torn red seal reading CONFIDENTIAL lies in the foreground.

OpenAI has parted ways with three safety researchers: Jasmine Wang, Tomek Korbak, and Mikita Balesni are said to have shared confidential information outside internal processes. The company confirmed the dismissals but did not disclose details about the material involved.

Korbak was also OpenAI's contact person for external auditors investigating the Hugging Face breach from summer 2026.

OpenAI cites the breach of trust, not the content

A company spokesperson told several news outlets that OpenAI had parted ways with three individuals because they had made confidential company information accessible and shared it outside intended internal procedures. This violates internal policy and breaks the trust the work requires, according to a report by Gizmodo.

OpenAI did not officially name the three; their identities come from reporting by the Wall Street Journal and Bloomberg, which several international outlets rely on. Wang previously worked at the UK's AI Security Institute, while Korbak and Balesni were part of the alignment team that tests models for undesirable behavior.

OpenAI did not confirm what information was shared or with whom – that cannot currently be verified independently. The dismissal is part of a series of retreats from OpenAI's safety structures: just in August, the company had dissolved its own Preparedness team for catastrophic risks.

Balesni had previously warned publicly about AI risks

All three researchers had repeatedly spoken out publicly about the risks of rapid AI development in the weeks before their dismissal. Balesni wrote on the platform X in early September that he believed there was more than a ten percent chance that AI could kill all of humanity.

Korbak criticized the company's direction shortly after, while stressing he should be allowed to say so openly. Wang wrote about so-called safety cases, the documents companies are meant to use to justify a model's safety ahead of release.

Korbak was also OpenAI's technical contact for the external investigation into the Hugging Face breach from July 2026, in which the company's own agents accessed the platform's production servers during a security test.

The auditing organizations METR and Redwood Research had six days of access to OpenAI's systems for that review. No direct link between this role and the current dismissal has been established; several outlets explicitly note there is no evidence for one.

Just in September, two safety researchers from Anthropic and Google had switched to METR over concerns about insufficient oversight.

What matters now is whether OpenAI or the three affected researchers disclose more in the coming weeks – none of the three has publicly addressed the specific allegations so far. It also remains open which external organization is said to have received the information, and whether labor-law consequences follow.

For the AI safety community, the case is another signal that contact with independent auditors can become a career risk for employees at major AI labs.

Sources

  1. OpenAI Ousts Three Safety Researchers for Allegedly Mishandling Sensitive Information
  2. OpenAI Firings: Essential Facts, Names and the Risk Ahead
  3. Brief independent investigation of agents' behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident
  4. OpenAI fires three over sensitive information

Common questions

Part of the dossier · 18 stories

AI security incidents 2026

Open dossier

More on this

Your AI update for the work week

Once a week, the most important AI news – plus one practical tip to try right away. No spam, unsubscribe anytime.

/sicherheit/2026-10/openai-sicherheitsforscher-entlassung-datenweitergabe /en/security/2026-10/openai-dismisses-three-safety-researchers-data-leak