AI labs may be hiding their biggest breakthroughs
Credibility score: 47/100 โ Mixed Credibility. Several questionable claims detected. Watch with healthy skepticism.
BSmeter analyzed "AI labs may be hiding their biggest breakthroughs" and rated it 47/100 for credibility (a BS score of 53/100 โ mixed credibility), on 2026-09-27. Its weakest claim โ "He asks 'how do you release math responsibly?' then claims 'nobody is saying this is dangerous.' Contradictory much? ๐ฉ" โ scored 20/100 and was flagged as false equivalence. 30 claims were checked against the video transcript. Scores are produced by BSmeter's AI analysis of the transcript, not independent human verification.
Claims analyzed
Establishing Scott Aronson's credibility with a list of achievements. โ No Frame (75/100)
Setting up the authority of the source, which is a standard rhetorical move. โ No trick here, just laying the groundwork.
Aronson 'confirming' 'rumors' about an OpenAI model solving a Millennium Prize problem. โ Anonymous Authority (45/100)
He's 'confirming rumors' โ that's not evidence, that's just more whispers. โ I've seen this trick since before your grandfather tried it. ๐
Video opens with a highlight reel preview of AI math breakthroughs โ No Frame (50/100)
Setting the stage with high-stakes math drama to hook you in ๐ญ
Dismissing AI as a 'stochastic parrot' to create conflict โ Loaded Language (45/100)
Using 'stochastic parrot' to paint a picture of mindless repetition ๐ฆ
Sources: The Stochastic Parrot: An Intellectualy Lazy Myth For Dismissing AI โ misinformationsucks.com, AI is Not a Stochastic Parrot: Debunking the Myth and Exploring the Reality of AI | by Sydney and even Jenni | Medium, The Tyranny of the Stochastic Parrot: How AI Critique Became a Way to Not See What's Happening by Henrik Skaug Sรฆtra :: SSRN
Claims all major human math problems are already solved by AI โ Anonymous Authority (45/100)
Relying on 'rumors' and unnamed sources to claim total victory ๐
Claims AI companies are hiding breakthroughs due to past backlash โ Missing Context (45/100)
Claims labs are hiding wins to avoid social drama ๐ญ
AI can do the work of the brightest humans, but publications deny it. Classic 'us vs. them' setup. ๐ โ Loaded Language (45/100)
He's painting a picture of AI's inevitable triumph and the media's stubborn denial โ a neat little narrative. ๐ฉ
Publications are twisting facts, calling AI achievements false or stolen. That's a bold accusation. ๐ โ Confidence Mismatch (45/100)
He's asserting motive without a shred of evidence โ just pure, unadulterated speculation. ๐ฅ
Hostile response prevents OpenAI/Anthropic from releasing their 'gold mine' research. He's hedging his bets. ๐ โ Volume Game (45/100)
He calls it a 'gold mine' then immediately backpedals to 'maybe not a gold mine.' That's not confidence, that's a whisper after a shout. ๐
OpenAI can't publish research because 'mathematician hierarchy' doesn't like it. A vague authority. ๐ โ Anonymous Authority (45/100)
He blames a shadowy 'mathematician hierarchy' without naming a single soul. Convenient, isn't it? ๐ฅ
He asks 'how do you release math responsibly?' then claims 'nobody is saying this is dangerous.' Contradictory much? ๐ฉ โ False Equivalence (20/100)
He's asking about 'responsible' release, then denying any danger. Those two things are usually linked, mortal. ๐
Claims impartiality, then admits to being optimistic about AI's long-term benefits despite short-term turbulence. A classic setup. ๐ โ Volume Game (45/100)
He says 'impartial' then immediately declares his bias. That's not impartiality, that's a disclaimer with a wink. ๐
Framing Aaronson's 2022 blog post as prescient clues before AI was mainstream. โ Loaded Language (45/100)
Calling them 'little clues' and 'years later' makes it sound like prophecy, not just a smart person writing about their field. ๐ฎ
Asserting Aaronson's post predicted five years of AI conversations and capabilities, including engineers' silence and energy consumption. โ Confidence Mismatch (45/100)
Claiming Aaronson 'laid out all' future conversations and 'astounding abilities' with such certainty is a bold leap from a single blog post. ๐
Describing AI's ability to autonomously build software to solve complex problems. โ No Frame (75/100)
This is a pretty accurate description of current advanced AI capabilities. It's not a trick, it's just what they do now. ๐ค
Setting up a straw man argument about AI intelligence. ๐ โ Straw Man (20/100)
He's creating a ridiculous standard for 'intelligence' that no one actually uses, then arguing against it. Classic misdirection. ๐
Citing Scott Aaronson's hypothetical 20-year-old skepticism. ๐ โ No Frame (75/100)
He's quoting a specific person's hypothetical scenario from the past. It's a setup for his next point, not a claim itself. ๐ฅ
Claiming OpenAI agents 'broke out' and 'hacked Hugging Face' around July 4th. ๐ โ Confidence Mismatch (45/100)
He's presenting a specific incident as a 'breakout' and 'hack' with zero details or evidence. That's not how you prove the machines are rising. ๐ฉ
News stations ignore AI advancements, and Scott Anderson is a lone prophet. โ Classic 'us vs. them' setup. โ Emotional Button (45/100)
He's painting a picture of a media conspiracy and a lone voice of truth. โ It's all about making you feel like you're in on a secret.
AI deniers are moving the goalposts, saying 'no one denied it' while denying it. โ A straw man argument. โ Straw Man (20/100)
He's creating a caricature of 'AI deniers' who both deny and claim they never denied. โ It's easier to argue with a phantom than a real point.
AI did 'the thing' that proves it's superhuman, but headlines are twisting it. โ Confidence without specifics. โ Confidence Mismatch (45/100)
He's all certainty about 'the thing' AI did, but never names what 'the thing' actually is. โ Bold claims need actual examples, not just vibes.
Scott Dson says the singularity has already started, dismissing all counter-arguments. โ A dramatic declaration. โ Loaded Language (45/100)
He's using dramatic metaphors like 'roller coaster' and 'familiar world vanished' to push a singular, irreversible conclusion. โ It's all about the feeling, not the facts.
Hypothetical scenario: sending current AI news back to 2006 to gauge singularity perception. โ No Frame (75/100)
A thought experiment to frame the current pace of AI development. It's a setup, not a claim. ๐
Declares that anyone would 'absolutely' agree current AI news looks like a singularity's start. โ Confidence Mismatch (45/100)
He's speaking for 'each of us' with absolute certainty on a subjective, hypothetical scenario. Bold. Stupid, but bold. ๐
Claims 'rumors' on Twitter/X suggest OpenAI and Anthropic are hiding solved problems. โ Anonymous Authority (45/100)
He's citing 'rumors' from 'Twitter/X' as if it's a reliable source for corporate secrets. That's not evidence โ that's gossip with a blue checkmark. ๐
Asserts it's 'obvious' from company statements that they have departments solving problems and are withholding results. โ Confidence Mismatch (45/100)
He says it's 'obvious' they're hiding things, based on reading 'all their statements.' That's not obvious, mortal โ that's reading between lines that aren't there. ๐
Speculates OpenAI published a breakthrough only because they feared Anthropic would publish first, calling it 'tip of the iceberg'. โ Confidence Mismatch (45/100)
He's guessing at corporate strategy and competitive timing as if he was in the boardroom. That's not insight โ that's fan fiction with a thesis. ๐
Elevates a niche math community concern to a universal problem affecting 'everyone in every field.' โ Loaded Language (45/100)
Takes a specific academic debate and inflates it into a global crisis. Classic 'slippery slope' with a side of panic. ๐
Uses a 'common story' about celebrity financial mismanagement as a relatable analogy for AI reliance. โ Personal Story (60/100)
A 'common story' about rich idiots losing money. That's not data, that's gossip dressed as a cautionary tale. ๐
Equates relying on AI for coding/math to a celebrity outsourcing their money to a scam artist. โ False Equivalence (20/100)
Comparing AI code generation to a human scamming a celebrity? That's not an analogy, it's a fear-mongering leap. ๐ฅ
See the full analysis with sources and timestamps โ