Read online version of the 9 October 2026 issue

Newsletter/9 October 2026

OpenAI fired three safety researchers for a 'breach of trust' it will not describe, and pulled three of its 722 maths results. 300 of 719 are computer-checked.

Both sides of the OpenAI firing in their own words - the four-page letter and the 06:17 UTC reply - with no document behind either; "OpenAI withdraws three mathematical results" against the repository's own history file; and the 12 November check, one prompt that tells you which of Anthropic's new rules touch your use of Claude.

By Daily Aletheia · Checked against the primary source · 9 October 2026 · 5 min read


Image: The researchers' letter "OpenAI cannot make AI safe on its own" - its first page as posted by Mikita Balesni on X, 8 October 2026 (mikitabalesni.com/letter/letter.pdf)

Last week OpenAI fired three of the people whose job was to watch what its models think. On Thursday they published a letter; early on Friday OpenAI answered it. In between, OpenAI withdrew three of the 722 mathematical results it announced on Monday.

The Signal - OpenAI fired three safety researchers and calls it "a significant breach of trust"; they say it was for talking to the auditors

What happened: Last week OpenAI fired three of its safety and alignment staff - Tomek Korbak, Jasmine Wang and Mikita Balesni. On Thursday they published a four-page letter, "OpenAI cannot make AI safe on its own". At 06:17 UTC on Friday OpenAI's newsroom account answered: "a thorough investigation found they violated clear policies on handling sensitive information" and uncovered "a significant breach of trust beyond what's outlined in the letter".

The details:

  • OpenAI's account: the decisions "were not about raising safety concerns or speaking out"; it names no policy and no act; contracts with "third-party safety assessors" come "in the coming weeks".
  • Korbak's account: he "was told verbally I was fired because of the way I communicated with METR" - the outside auditor of the summer's Hugging Face breach. "Talking to METR was my job."
  • The letter: the three "were not the source of the leak for The Information article" about less monitorable architectures; Wang's access to an executive's email "was delegated for recruiting purposes, with permission", and when she asked for it to be removed "IT failed to do so".
  • Where both sides agree: the letter states "The monitorability of frontier models is degrading" and quotes Jakub Pachocki, OpenAI's chief scientist, calling it "fragile and unfortunately trending in a negative direction"; OpenAI's note agrees.

What it means for you: OpenAI's point of contact with its outside auditor is gone; the auditor contracts it says it is finalising are the thing to watch. The letter cites Sam Altman's 12 September commitment "to give independent evaluators ongoing, employee-like access" - read the contracts for that line.

The counterweight: Neither side has produced a document: OpenAI names no policy, date or act; the researchers' account of their own conduct is their own.

Correction: issue 007's Signal gave Mistral Large 4's active parameters as 49 billion, the figure on Mistral's page on 6 and 7 October; Mistral changed its page to 52 billion on 8 October, and the issue and the story now carry that with a dated note.

The Truth Check - "OpenAI withdraws three mathematical results"

Free to keep reading. Every issue, checked, three mornings a week.

Free. Unsubscribe any time.

The Move - The 12 November check

The Move - The 12 November check is for members. $9 a month, founding price.

The Check

What was said
"OpenAI withdraws three mathematical results" - the Hacker News headline (323 points, 578 comments by Friday noon) on Dan Roberts's post of 8 October, read as OpenAI's 722 results falling apart
What the Source Says
history.md, dated 7 October: "a sign error invalidates a stabilization-trace cancellation argument and the construction used by two dependent papers" - three manuscripts withdrawn and named; fourteen revised with "proof repairs, corrected statements, clearer hypotheses and dependencies"; six formalisations added; the README now reads "719 manuscripts organized into 372 families"; "300 / 719 = ~42%" of top-line results formalised.
The Difference
The withdrawal count holds exactly and OpenAI published it itself. The number under it is the one to carry: 419 of the 719 remaining results have no computer-checked proof, and the first outside specialist to read one (Karagila, the Partition Principle paper) calls it unreadable as written - not wrong, unrefereed.

Sources & further reading

  1. 01OpenAI Newsroom on X - A note from our research leaders (9 Oct 2026, 06:17 UTC) ↗Primary
  2. 02Korbak, Wang and Balesni - OpenAI cannot make AI safe on its own (letter, 8 Oct 2026) ↗Primary
  3. 03Tomek Korbak on X (8 Oct 2026) ↗Primary
  4. 04Mikita Balesni on X (8 Oct 2026) ↗Primary
  5. 05Jasmine Wang on X (8 Oct 2026) ↗Primary
  6. 06TechCrunch - Fired OpenAI safety researchers dispute misconduct claims, warn of chilling effect (8 Oct 2026) ↗Attributed
  7. 07OpenAI - the openai/math repository, history.md and README (7-8 Oct 2026) ↗Primary
  8. 08Dan Roberts (OpenAI) on X (8 Oct 2026) ↗Primary
  9. 09Hacker News - the thread on the withdrawals, read 9 Oct 2026 ↗Attributed
  10. 10Asaf Karagila - OpenAI, the Partition Principle, and mathematics (8 Oct 2026) ↗Primary
  11. 11Álvaro Lozano-Robledo, guest post on Terence Tao's blog - What should we tell our students? (8 Oct 2026) ↗Primary
  12. 12Anthropic - 2026 Usage Policy update (8 Oct 2026) ↗Primary
  13. 13USA Today Co., Inc. v. OpenAI Foundation, complaint, S.D.N.Y. 1:26-cv-08892 (8 Oct 2026) ↗Primary
  14. 14Anthropic - Introducing the Anthropic Cyber Mission (8 Oct 2026) ↗Primary
  15. 15Anthropic - Building on our commitment to American scientific discovery (8 Oct 2026) ↗Primary
  16. 16Google Cloud - Welcome to Gemini at Work 2026: Introducing the Gemini agent (8 Oct 2026) ↗Primary

Was this issue useful?