The controversy around OpenAI "solving" a $1m math problem (Navier-Stokes)
Credibility score: 62/100 — Mostly Credible. Mixed credibility - some claims are solid, others need verification.
BSmeter analyzed "The controversy around OpenAI "solving" a $1m math problem (Navier-Stokes)" and rated it 62/100 for credibility (a BS score of 38/100 — mostly credible), on 2026-09-10. Its weakest claim — "Claims sub-problems equal Navier-Stokes solution — false equivalence" — scored 20/100 and was flagged as false equivalence. 30 claims were checked against the video transcript. Scores are produced by BSmeter's AI analysis of the transcript, not independent human verification.
Claims analyzed
OpenAI solved Navier-Stokes today — framing the claim as breaking news — Confidence Mismatch (45/100)
Announces a solved $1M problem on the exact day of the video — no source, no paper, just the date as proof. 😈
Accusations of credit pressure + technical objection still pending — No Frame (75/100)
Names the actual mathematicians and dates — receipts on the table.
Names two mathematicians accusing OpenAI — credit pressure and proof objection — No Frame (75/100)
Straight reporting — names the people and what they said.
Only specialists can judge the math — rest of us can't tell — No Frame (75/100)
Honest admission that the real verdict is above his pay grade.
Admits only specialists can judge the proof — hedges his own authority — No Frame (75/100)
Calls his own limit before he even starts — rare honesty.
Navier-Stokes = $1M unsolved problem, only one solved so far — No Frame (75/100)
Correct setup: stakes and scarcity are real.
Perelman rejected prize and medal over credit dispute — No Frame (75/100)
Accurate historical parallel — credit fights aren't new here.
Perelman refused prize and medal over credit dispute — historical parallel — No Frame (75/100)
Uses real precedent to frame the current credit fight.
Aug 15: Buckmaster + Anthropic guy prove two related equations — No Frame (75/100)
Timeline and names locked in — zero ambiguity.
Dates the related proofs to August 15th — sets timeline before OpenAI's claim — No Frame (75/100)
Pins the timeline with specific names and dates.
Dates a Lean verification to August 22nd, 2026. — No Frame (75/100)
Straight date, no spin — just putting the timeline on record.
Presents OpenAI’s 100-page proof as fact — missing context on verification — Missing Context (45/100)
No word on whether the proof passed peer review or even Lean check — just “they said they have it.”
OpenAI produced a 100-page Navier-Stokes blow-up proof — per Buckmaster. — Missing Context (45/100)
The 100-page proof exists only in Buckmaster's retelling — OpenAI hasn't released it.
Claims sub-problems equal Navier-Stokes solution — false equivalence — False Equivalence (20/100)
Blow-up in three related equations ≠ solved the actual prize problem. Different beasts.
Buckmaster released three blow-up papers plus a statement accusing OpenAI of credit manipulation. — No Frame (75/100)
Speaker lays out the public record without embellishment.
Leaves credit fight at “he said / they said” — confidence mismatch — Confidence Mismatch (45/100)
Speaker states both sides’ positions without evidence either way. Feels resolved when it isn’t.
Claims OpenAI solved Navier-Stokes via 10k agents in 88 hours — Confidence Mismatch (45/100)
Says 'proof' like the math community already signed off — they didn't. 🔥
Announces 'proof' from 10k agents in 88 hours — confidence mismatch — Confidence Mismatch (45/100)
Calls it a 'proof' when it's an unreleased internal model — that's a finished product wearing a lab coat.
Lean verification validates the proof — missing context — Missing Context (45/100)
Lean checks syntax, not whether the axioms were right to begin with. That's like proofreading a novel for typos and calling it canon.
Palasek found a 'weak point' today — cherry-picked framing — Cherry-Picked (45/100)
Singular 'weak point' implies one fixable flaw. The objection actually threatens the mechanism that lets the instability grow at all.
Presents Palasek's objection as a fatal flaw — Missing Context (45/100)
Treats one mathematician's objection as game-over — ignores that Tao already suggested next steps. 🍒
Tao 'welcomed' the objection — loaded language — Loaded Language (45/100)
'Welcoming' sounds collegial. Tao is publicly signaling the proof isn't settled, which is mathematician for 'this needs work.'
Correctly states Clay Prize rules after the fact — No Frame (75/100)
Finally drops the actual rules everyone forgot — prize isn't even on the table yet. ✅
Even if fixed, two-year wait for prize — no frame — No Frame (75/100)
Straight rules citation. The Clay Institute's two-year cooling-off period is real, boring, and correctly stated.
Invokes 'history' as blanket doubt without naming any example. — Anonymous Authority (45/100)
Historical counter-examples exist, but none are cited — just 'many things.'
Prize needs community acceptance — true, but framing hides the real fight — No Frame (75/100)
Correct gatekeeping — no one's handing over the million until the math checks out.
Past proofs looked solid then fell apart — hedging with 'whatever' while making a real point — No Frame (75/100)
He's right that math history is littered with 'done deals' that later cracked. The 'whatever' is just verbal filler.
OpenAI says it's not chasing the prize — just flexing the model — No Frame (75/100)
They're explicitly not claiming the million. The move is to treat the math as a capability demo instead.
A million dollars is pocket change to OpenAI — true, but misses the actual stakes — Missing Context (45/100)
The money's irrelevant. The prestige of breaking a Clay problem is the real currency here.
Four 'we'll see' hedges in a row — zero conclusions drawn. — No Frame (75/100)
Straightforward: the situation is still unfolding and the speaker admits it.
See the full analysis with sources and timestamps →