Ilya Sutskever new "Superintelligence" model will change EVERYTHING
Credibility score: 42/100 — Mixed Credibility. Several questionable claims detected. Watch with healthy skepticism.
BSmeter analyzed "Ilya Sutskever new "Superintelligence" model will change EVERYTHING" and rated it 42/100 for credibility (a BS score of 58/100 — mixed credibility), on 2026-08-24. Its weakest claim — "Uses real NSA warning to prove AI superintelligence is an 'active threat'" — scored 20/100 and was flagged as false equivalence. 30 claims were checked against the video transcript. Scores are produced by BSmeter's AI analysis of the transcript, not independent human verification.
Claims analyzed
Gemini 4 'finished pre-training' — zero sources named — Anonymous Authority (45/100)
Says 'we already have proof' — names none. Classic empty citation.
Names three upcoming models without naming sources — Anonymous Authority (45/100)
Gemini 4, Chinese model, SSI debut — drops three bombs, cites zero. Classic anonymous authority.
Promises multiple 'huge' models by year-end with total certainty — Confidence Mismatch (45/100)
From 'might' to 'will' with zero evidence in between. Bold. Stupid, but bold.
Claims Mythos 5 was the first model refused for safety reasons — Missing Context (45/100)
Treats a marketing line as historical fact. I've watched labs say 'too dangerous' since before your grandfather tried it.
Uses real NSA warning to prove AI superintelligence is an 'active threat' — False Equivalence (20/100)
NSA warns about AI-assisted hacking today — speaker leaps to 'superintelligence is here and dangerous.' That's not a link, mortal. That's a magic trick.
Cites Iranian hack of UK plant as evidence of current AI-powered attacks — Missing Context (45/100)
Top comment already corrected this: small gas generator, happened last month, no AI mentioned. Story's doing the heavy lifting here.
Admits 'nothing to compare it to' yet still ranks it #1 — self-own — Confidence Mismatch (45/100)
Own words kill the ranking — if nothing exists to compare, the crown is fiction 🔥
Admits no AI link in reports — still floats the possibility anyway. — Volume Game (45/100)
Says 'no mention of AI' then spends the next minute implying AI anyway. Classic volume play. 💀
Quotes NSA/FBI as uncertain on Iran link — uses their doubt to build case. — Missing Context (45/100)
Cites agencies saying they're unsure — then treats that uncertainty as supporting evidence. 😈
Flags official uncertainty on Iran link — then keeps the headline anyway — Missing Context (45/100)
Tells us the agencies aren't sure, then keeps selling the Iran story like it still holds 🚩
NSA quote on AI lowering attack barriers — applies it to this incident without evidence. — Missing Context (45/100)
NSA warns AI lowers the bar in general — speaker pins it to this specific attack with zero proof. 🔥
Quotes agencies on AI lowering attack barriers — but never ties it to this incident — Missing Context (45/100)
Agencies warned about AI scripts in general — he weaponizes the quote for this specific plant hit 😈
Frames skepticism as gullibility — 'power plant to sell you' jab. — Emotional Button (45/100)
Doubters get mocked as suckers — emotional pressure to agree AI was involved. 💀
Jokes 'power plant for sale' if you doubt AI involvement — pressure tactic — Emotional Button (45/100)
Mockery replaces evidence — 'believe me or you're a mark' energy 😈
Calls UK plant hack 'most successful ever' — no comparison offered. — Confidence Mismatch (45/100)
Labels it the biggest UK cyberattack — then admits nothing exists to compare it to. 😈
Names 'Fable 5' as first truly dangerous model — no source, no receipts — Anonymous Authority (45/100)
Drops a model name like gospel — zero paper, zero benchmark, just the word 'truly' doing heavy lifting 💀
Lumps Bitcoin losses and plant shutdown under one AI umbrella — two different events, one assumption — False Equivalence (20/100)
Two unrelated incidents, one causal leap — he's gluing headlines together with hope 😈
AI agents now dominate token usage — humans no longer matter most — Missing Context (45/100)
Assumes the chart proves agents are doing the hacking — chart only shows volume, not intent or success
Chart shows AI agents already dominate token usage — Anonymous Authority (45/100)
'If you look at the chart' — no chart shown, no source named. Trust me, bro.
NSA/FBI report: AI agents scanning exposed industrial PLCs right now — Missing Context (45/100)
Cites a specific agency report — never links or names the document. The devil is in the missing footnote.
Agencies only warn about reconnaissance, not attacks yet — No Frame (75/100)
This part is straight: agencies flagged scanning, not yet active exploitation.
Agencies only see reconnaissance, not active attacks yet — No Frame (75/100)
Accurately reflects the report's distinction between scanning and striking
Proof AI agents succeed: good guys also find bugs at scale — False Equivalence (20/100)
Conflates human security researchers with autonomous AI agents — two very different toolkits.
Agents 24/7 with infinite clones — history's first time — Confidence Mismatch (45/100)
Says 'never before in history' like automation is brand new. Script kiddies have had bots for decades.
Jumps from 'might' to coordinated multi-industry attack with zero evidence — Confidence Mismatch (45/100)
From possibility to synchronized collapse in one sentence — zero probability data between them 😈
One exploit wave could crash banks and power grids — Emotional Button (45/100)
Paints simultaneous multi-industry collapse as the default outcome. Fear does the heavy lifting.
Cascading financial collapse scenario presented as realistic outcome — Emotional Button (45/100)
Triggers 2008-style panic without showing how AI agents would coordinate the timing or scale needed
NSA/FBI claimed only elite humans could do cyber ops — now AI changes that — No Frame (75/100)
Straight summary of what the agencies actually said before AI agents arrived
NSA/FBI claim: only elite hackers can do real cyber ops — Missing Context (45/100)
States what the agencies 'said before' — no quote, no date, no source. Classic anonymous authority.
Dismisses skeptics as wrong without engaging their actual arguments — Straw Man (20/100)
Reduces all skepticism to 'PR stunt' dismissal — ignores legitimate questions about timeline and capability claims
See the full analysis with sources and timestamps →