OpenAI's secret model just BROKE math...

Credibility score: 49/100 β€” Mixed Credibility. Several questionable claims detected. Watch with healthy skepticism.

BSmeter analyzed "OpenAI's secret model just BROKE math..." and rated it 49/100 for credibility (a BS score of 51/100 β€” mixed credibility), on 2026-10-07. Its weakest claim β€” "AI will either help us or leave us in the dust like monkeys β€” False Dilemma" β€” scored 20/100 and was flagged as false dilemma. 26 claims were checked against the video transcript. Scores are produced by BSmeter's AI analysis of the transcript, not independent human verification.

Of 26 claims analyzed: 1 scored under 40, 22 between 40 and 69, and 3 at 70 or above.

Claims analyzed

Claims OpenAI's secret model solved massive math problems via hype-driven questions β€” Confidence Mismatch (45/100)

At 0:00

Asking if it 'solved math forever' is a massive reach. That's not news, that's just asking for attention πŸ’€

Why this score: The speaker uses rhetorical questions about solving the Riemann hypothesis and P vs NP to manufacture a sense of world-changing importance before even explaining what actually happened.

Original quote: β€œWell, it looks like the rumors were true. Openai just published 722 mathematical manuscripts, all of them produced by this unreleased insider internal model. Did it solve the Reman hypothesis? Did it prove that P equals NP? Did it solve whole math forever?”

Labels the release as 'absolutely insane' to manipulate emotion β€” Loaded Language (45/100)

At 0:20

'Absolutely insane' is doing all the heavy lifting here. Pure emotional bait 🚩

Why this score: The speaker pivots from the big questions to a vague 'insane' descriptor, using high-intensity adjectives to mask the fact that they haven't actually provided a concrete result yet.

Original quote: β€œWell, not quite. But what it released is absolutely insane. Some people refer to it as the quasi reman hypothesis. to the people that have no idea what any of that stuff is.”

OpenAI's Navier-Stokes win is a technicality, not a claim of victory β€” Missing Context (45/100)

At 0:47

He's glossing over the massive drama and credit fights happening behind that 'not claiming it' line 🎭

Why this score: He frames the lack of a prize claim as a polite choice, ignoring that mathematicians are currently accusing them of stealing work and causing an absolute scandal.

Original quote: β€œOpenai already published Navy Stokes, which sounds like it would qualify for that prize. Now, OpenAI is not claiming it because it was solved by an AI model.”

Math output is increasing by 10x every single month β€” No Frame (75/100)

At 1:00

The numbers are actually insane. No trickery here, just pure industrial-scale math πŸ“ˆ

Why this score: The data is verified by current events; the jump from 10 to 100 to 722 is a massive, verifiable spike in output.

Original quote: β€œIn August of this year, about two months ago, we published something like 10 results on 10 problems. In the next month, September, right, that's when they published Navy Stokes plus 100 plus other math problems. And today we're 6 days into the month and they publish 722 manuscripts”

claims news outlets are twisting the truth about stolen work 🚩 β€” Missing Context (45/100)

At 2:59

He's glossing over the actual dramaβ€”mathematicians are genuinely accusing them of theft 🎭

Why this score: He frames the 'stealing' narrative as a media twist, but web context shows it's an active, explosive scandal involving real credit disputes.

Original quote: β€œthat the AI model found a lot of news publications try to twist it as in they stole it. Like there was this mathematician who did all the work and opening eye they stole it from him.”

claims a general LLM can solve complex math via simple tasks 🚩 β€” False Equivalence (45/100)

At 4:03

Comparing writing a poem to solving frontier math is wild. That's not the same league πŸ’€

Why this score: He equates low-level generative tasks like writing emails to high-level mathematical discovery, masking the massive complexity gap.

Original quote: β€œThis isn't even like a math model. This is a large language model. They can write a rough draft of your email, write you a poem, do some little coding, and also progress humanity's understanding of mathematics”

Explaining the Riemann hypothesis with zero technical authority 🎭 β€” Confidence Mismatch (45/100)

At 4:37

Admitting they're lost while trying to explain high-level math with zero credentials πŸ’€

Why this score: The speaker uses a 'learning in real-time' vibe to mask the fact that they are just reciting definitions without actual mathematical depth.

Original quote: β€œI feel like I understand less. But we have what people have called the quasi reman hypothesis. So the reman hypothesis is basically looking at how prime numbers behave.”

Using a mountain metaphor to simplify complex math boundaries πŸ”οΈ β€” Missing Context (45/100)

At 6:03

Using a mountain and fog metaphor to gloss over the actual complexity of the math 🌫️

Why this score: The metaphor simplifies the 'boundary' concept so much it risks losing the actual precision required for a mathematical proof.

Original quote: β€œSo AI established a fixed 0.875 boundary. Reman's target is 0.5 and humans we got to the 1.0 line.”

Claims a 'big step forward' despite only partial solutions 🚩 β€” Confidence Mismatch (45/100)

At 6:35

He's acting like a partial win is a massive victory. It's just one piece of the puzzle, not the whole damn thing πŸ’€

Why this score: The speaker acknowledges these cases don't settle the entire problem, yet immediately pivots to calling it a 'big breakthrough.' It's an attempt to manufacture hype from incomplete data.

Original quote: β€œAnd again, we have a $1 million prize for the full set of these problems. So these cases don't settle the entire problem, but again, it seems like it's a big step forward.”

Predicts a global 'scramble' to verify AI discoveries ⚠️ β€” Loaded Language (45/100)

At 8:14

Using words like 'flood' and 'scramble' to make a technical bottleneck sound like an apocalypse 🌊

Why this score: The speaker uses high-intensity language to describe the logistical challenge of peer review. It frames a massive workload as an unstoppable natural disaster rather than a systemic problem to be solved.

Original quote: β€œWhat happens when these AI models start spitting out a flood of discoveries in any given field? We're seeing it happen live to mathematics. And step one is going to be this scramble to figure out how do we verify these results?”

Claims AI results will rewrite textbooks via massive shifts in math education β€” Confidence Mismatch (45/100)

At 8:37

Speculating on a total textbook rewrite based on unverified model outputs πŸ“šπŸ€”

Why this score: The speaker is making massive, sweeping claims about the future of academia based on a handful of unreleased model results. It's high-stakes speculation without any actual educational implementation data.

Original quote: β€œthose are very different stories. A lot of the stuff here rewrites textbooks, right? It completely changes how we explain math in colleges and universities.”

Predicts big math problems will be solved soon with uncertain utility β€” Missing Context (45/100)

At 9:58

Predicts a solution is 'near' without defining what solving it actually looks like πŸ”πŸš©

Why this score: He's talking about 'encirclement' and getting closer to big problems, but doesn't address the massive controversy regarding credit theft or how these 'solutions' are even being verified by humans.

Original quote: β€œit does feel like soon that will be solved as as well. But the big question is like what actionable things will flow out of this?”

AI will either help us or leave us in the dust like monkeys β€” False Dilemma β€” False Dilemma (20/100)

At 13:05

He's cornering himself into two extreme outcomes. It's a binary trap for a non-binary future πŸ’€

Why this score: The speaker presents a false dichotomy: we either gain total understanding or become 'monkeys' unable to comprehend our own tech. It ignores the messy middle ground of partial understanding or collaborative evolution.

Original quote: β€œwill it improve our understanding or will be completely lost just like the monkeys. So right now for a lot of that stuff there's not that many people in the world that truly understand it”

OpenAI has massive math breakthroughs β€” Anonymous Authority β€” Anonymous Authority (45/100)

At 14:19

Relying on 'rumors' and a single blog post to validate massive breakthroughs. Weak 🚩

Why this score: He's using 'rumors' and a specific blog post as his only foundation for claiming massive breakthroughs. While the web context confirms OpenAI released 722 manuscripts on Oct 6, he's framing it as a 'rumor' rather than citing the actual massive release of data.

Original quote: β€œthese rumors are true. Opening eye is sitting some on some pretty big mathematical breakthroughs.”

Speculative question about preserving human mathematical expertise β€” Missing Context (45/100)

At 14:30

Framing a massive industrial shift as a philosophical choice about 'human community' 🎭

Why this score: He's presenting a massive structural shift in how math is done as a simple question of 'do we want to keep people around.' It ignores the economic and structural pressures driving it.

Original quote: β€œBut the point is he asked this question as far as mathematicians go. Do we want is is it our goal to maintain this human community across generations of people that are capable of understanding the sort of frontier math?”

Vague description of historical math progress β€” Loaded Language (45/100)

At 14:56

Using 'pushed the field forward' to romanticize old-school research πŸ’€

Why this score: It's high-level fluff. He's using loaded, positive descriptors to create a 'human vs machine' conflict without actually discussing the mechanics of progress.

Original quote: β€œBut that group of people, they sort of understood the latest math available. They worked on it and they they pushed the field forward moving forward.”

Speculative claim that math automation boosts human biology β€” Confidence Mismatch (45/100)

At 15:33

Leaps from math automation to 'longer human lifespans' without a bridge 🚩

Why this score: He's making massive, unearned leaps of logic. There is no direct line from an AI solving equations to a breakthrough in human biology that isn't just pure optimism.

Original quote: β€œBut on the flip side of the coin, what if this sort of industrialization of math, what if it leads to breakthroughs that improve, let's say, human lifespan or our understanding of how our brains work”

Comparing a future without work to being 'house cats' β€” Just Vibes (50/100)

At 16:05

Using a 'house cat' analogy to sell a utopian future 🐱

Why this score: This is pure rhetorical flavor. It's a vibe-based comparison to make the idea of post-work existence sound cozy rather than potentially existential or boring.

Original quote: β€œthey're just kind of just having a good old time and they don't really have to worry about anything. They're like really well taken care of house cats.”

Claiming math is the first industry to be disrupted due to verifiability β€” No Frame (75/100)

At 16:26

Actually makes a solid point about the nature of math being verifiable πŸ›‘οΈ

Why this score: This is his strongest point. It's a logical observation about why math is the testing ground for AI autonomy, backed by current trends in model output.

Original quote: β€œIt's happening here in mathematics first because it's verifiable among other things.”

Predicting AI will solve all human ailments via 722 medicines β€” extreme extrapolation β€” False Equivalence (45/100)

At 16:36

Using a hypothetical 'fix all ailments' scenario to mask the uncertainty of biological complexity πŸ’‰πŸ’€

Why this score: He's equating mathematical breakthroughs to solving every biological ailment ever known. It's a massive leap in logic that ignores the messy reality of biology compared to pure math.

Original quote: β€œNext year when some AI model spits out 722 medicines that fix all the ailments that any human being has ever had, we're going to be asking a lot of the same questions.”

Claiming mathematicians are unhappy due to 25 field medalists signing a letter β€” No Frame (75/100)

At 17:33

Actually cites specific numbers and a group of experts. No tricks here, just the facts πŸ“œπŸ”₯

Why this score: He's citing a specific event involving 25 field medalists. While the 'unhappiness' is an interpretation, the existence of the letter and the specific group is verifiable.

Original quote: β€œa lot of mathematicians are not very happy with OpenAI right now. 25 field medalists have signed a letter.”

Claims 'everyone' wants a pill for better days based on casual conversation β€” Anonymous Authority (45/100)

At 18:43

Says 'everyone is okay with it' based on just 'talking to people' β€” no data, just vibes πŸ’€

Why this score: He’s using anecdotal social feedback as a universal truth. 'Talking to people' is the weakest form of data collection known to man.

Original quote: β€œEveryone, pretty much everyone is okay with the following. If you imagine a day where you were at your best... would you take a pill or do some treatment where you would just have a lot more of those days? ... What I found is most people would say yes to that.”

Claims 90% accuracy would be the biggest math event ever β€” Confidence Mismatch (45/100)

At 20:13

Hyping up a '90% accuracy' scenario as the biggest thing in history. Pure hype 🚩

Why this score: He's pivoting from the uncertainty of 'is it 90% or 100%' to a massive, hyperbolic conclusion. It's high-stakes speculation masquerading as a fact.

Original quote: β€œif this is like 90 plus% correct, this is the biggest thing in the history of math, I think, ever.”

Using a math professor's shock to set up a joke about AI 🎭 β€” Loaded Language (45/100)

At 20:30

Using 'shocked' and 'stunned' to build emotional weight for a punchline 🎭

Why this score: He's using the heavy emotional weight of a professor's realization to set up a comedic pivot. It's not about the math; it's about the drama of the setup.

Original quote: β€œWhen the plumber handed him the bill after an hour or so of work, my friend was stunned. He was a mathematics professor.”

Narrative pivot used to mask the real point of the video 🚩 β€” Missing Context (45/100)

At 20:56

Building a story about career changes to hide the actual AI math reveal 🚩

Why this score: The story about a professor becoming a plumber is just flavor text to distract from the actual discussion of AI models breaking math. It's all misdirection.

Original quote: β€œHe started making a lot more money. He started really enjoying life until one one day the company required everyone who worked there to join classes to finish 8th grade.”

A punchline used to bypass the actual technical explanation πŸ’€ β€” Just Vibes (50/100)

At 21:33

Ending on a joke instead of actual data about the model's failure πŸ’€

Why this score: He's pivoting from a math joke to an abrupt exit. It's pure entertainment value meant to mask the fact that we haven't actually discussed the model's specific 'breakdown' yet.

Original quote: β€œThey're going, "Switch the limits of the integral." So, on that note, I think I'll leave it here.”

See the full analysis with timestamps β†’