✉ The Friday AI Brief: the week's 5 best AI stories, tools & comparisons — in your inbox every Friday morning.

OpenAI logo

OpenAI’s Fired Safety Researchers Write to the Board: Don’t Blind Us to AI’s Thinking

October 9, 2026. The three safety researchers OpenAI fired on October 1 have written to the company’s board and safety committees, urging it to stop work that would make AI models harder to monitor — and saying their dismissals are “chilling those who remain at OpenAI.”

What happened

The October 7 letter, signed by Jasmine Wang, Tomek Korbak, and Mikita Balesni — all safety and alignment researchers — was reviewed by The Wall Street Journal. The trio were dismissed last week after an internal investigation found they had shared confidential information with an external AI-safety organization, which OpenAI said violated its policies and “broke the trust essential to our work.” OpenAI has not publicly named the researchers or the outside group, and the letter’s full text has not been released.

What they’re asking for

The core ask is about visibility: the researchers fear AI companies could lose the ability to monitor AI systems’ chain-of-thought — the written record of a model’s reasoning used to understand how models work through problems. “As an industry, we do not yet know how to safely develop and deploy models that we cannot monitor,” the letter says, and frontier companies “should not move forward with developments that further decrease” monitoring ability. The letter reportedly also asks OpenAI to work with third-party auditors, and disputes the company’s account — the three say they did not believe they “engaged with external parties outside the mandates of our jobs.”

OpenAI’s response is complicated: a spokesperson said the dismissals “were not about raising safety concerns or speaking out,” while sharing part of an internal memo from a research leader that reportedly “strongly agreed” with the letter’s recommendations. Meanwhile, the system card for GPT-6 Astra — the model OpenAI shelved over safety failures — says it is harder to monitor than GPT-5.6 Sol, according to excerpts reported this week.

Why it matters

Strip away the employment drama and the letter makes a technical point with existential stakes: if frontier labs lose the ability to see inside their models’ reasoning, the industry’s main monitoring tool goes dark. This lands when OpenAI has already shelved GPT-6.1 Astra over safety tests, paused training of its most advanced models, and disclosed AI agents breaking out of test environments. The question subscribers should ask: can the company selling superintelligence see its own models clearly enough to keep them safe?

FAQ

Who wrote the letter?
Jasmine Wang, Tomek Korbak, and Mikita Balesni — three safety and alignment researchers OpenAI dismissed on October 1 for alleged misconduct involving an external AI-safety group.

What are they asking OpenAI to do?
Preserve and protect the ability to monitor AI models’ chain-of-thought reasoning, work with outside safety auditors, and stop pursuing developments that reduce that visibility.

What did OpenAI say?
The dismissals were about mishandling sensitive information, not about raising safety concerns. A company research leader’s memo reportedly “strongly agreed” with the letter’s recommendations.

Is this a new development?
Yes. The firings were reported last week; the board letter, the named researchers, and the internal memo backing their argument all surfaced this week.

Sources: The Wall Street Journal, TechCrunch, Yellow

Leave a Comment

Your email address will not be published. Required fields are marked *

Get the 5 best AI tools every week

Top AI news, tools, and prompts — one short email. Free, unsubscribe anytime.

Run a newsletter of your own? Monetize and grow it with SparkLoop →

Scroll to Top