We need to talk about Jev...
Credibility score: 52/100 — Mixed Credibility. Several questionable claims detected. Watch with healthy skepticism.
BSmeter analyzed "We need to talk about Jev..." and rated it 52/100 for credibility (a BS score of 48/100 — mixed credibility), on 2026-09-18. Its weakest claim — "Claiming Jev is 'hundreds of times faster' than traditional LLMs, co-invented by a ChatGPT creator." — scored 45/100 and was flagged as confidence mismatch. 28 claims were checked against the video transcript. Scores are produced by BSmeter's AI analysis of the transcript, not independent human verification.
Claims analyzed
Introducing Jev as a new, incredibly fast AI — showing a demo. — No Frame (75/100)
They're showing a demo and claiming it's real-time. If it's not sped up, then it's just a demo. No trickery here.
Claiming Jev is 'hundreds of times faster' than traditional LLMs, co-invented by a ChatGPT creator. — Confidence Mismatch (45/100)
Hundreds of times faster? That's a hell of a leap, and they just glossed over it like it's a footnote. Where's the data, mortal? 💀
Claims Jev is 'hundreds of times faster' than traditional LLMs, with 'insane' speed. — Loaded Language (45/100)
Hundreds of times faster? 'Insane' speed? That's not a metric, that's a hype machine in overdrive. Show me the benchmarks, mortal. 😈
States Jev is 'so efficient' it's free for output tokens and 'fractions of a penny' for input. — Confidence Mismatch (45/100)
Free output and fractions of a penny for input? Sounds like a sales pitch dressed as a technical marvel. The 'efficiency' is just the hook. 💰
Claims Jev is 'up to 200 times faster' and 'up to 400 times cheaper,' with free output tokens. — Volume Game (45/100)
Now it's 'up to' 200x faster and 'up to' 400x cheaper. That 'up to' is doing a lot of heavy lifting to cover their ass. 🚩
Presents a benchmark showing Jev 'on par' with Luna, Terra, Sonnet 5, and above Opus 5 and Soul, at a 'fraction of a penny.' — Missing Context (45/100)
They flash a benchmark, name-drop some big models, and claim 'on par' or 'above.' But 'type safe' for what? What's the actual metric? Details, mortal, details. 💀
Claims Jev is 'so fast' it can play Doom in real time, making all decisions within the game loop. — Confidence Mismatch (45/100)
Playing Doom in real time? That's a hell of a demo, but 'making all the decisions' is a broad claim. Is it just reacting, or strategizing? Big difference. 🔥
Demonstrates Jev's speed in a 'wiki race' — pure spectacle, no real comparison. — Just Vibes (50/100)
They're showing off how fast it is, but without a human baseline, it's just a light show. 😈
Claims 'five hops in half a second' compared to other models taking 4-5 seconds — a direct speed comparison. — No Frame (75/100)
A direct comparison with specific numbers. Finally, something tangible. 😈
Describes Jev as a 'decision engine' capable of 'hundreds or thousands of decisions in parallel' — a bold claim of capability. — Confidence Mismatch (45/100)
Hundreds or thousands in parallel? That's a hell of a leap from a wiki race. Show, don't just tell. 😈
Presents a support ticket routing example, claiming Jev processes it in 'milliseconds' — another speed claim without direct proof. — Confidence Mismatch (45/100)
Another 'milliseconds' claim for a complex task. We're just supposed to take their word for it now? 💀
States Jev's motto is 'building prod not God,' implying a jab at Anthropic — a competitive framing. — Loaded Language (45/100)
Oh, a little corporate drama. 'Building prod not God' is a nice dig, but it's just marketing. 😈
Sets up 'catastrophic' hallucinations to elevate Jev's solution. — Emotional Button (45/100)
He's painting a picture of doom and gloom with 'catastrophic' to make Jev sound like the only savior. Classic fear-mongering. 💀
Claims Jev has 'zero hallucinations' after listing high-stakes scenarios. — Confidence Mismatch (45/100)
Zero hallucinations? In 'military targeting' and 'healthcare'? That's a bold claim for a new tech, especially without a shred of proof. 🔥
Zapier sponsor read, claiming '100x 200x the speed' and 'fraction of the price'. — Sponsored (50/100)
Ah, the old '100x speed, fraction of the price' magic trick. It's a sponsor read, mortals. Don't fall for the hyperbole. 😈
Clarifies Jev's role in the demo — not for building from scratch, but for decision-making. No trick here. — No Frame (75/100)
He's setting the stage for what Jev actually does, which is fair enough. No smoke and mirrors yet.
Presents specific numbers for AI character reactions to a 'fire sale' prompt. It's just a demo. — No Frame (75/100)
He's just showing the results of his little simulation. It's a demo, not a grand claim about AI behavior.
Describes AI characters' reactions to a 'poisonous snake' threat, noting some defiance. It's a demo, not a deep insight. — No Frame (75/100)
He's just showing the AI's response to a new prompt. It's a demonstration of its programmed behavior, nothing more.
Describes Jev powering chopsticks to sort 150,000 Skittles, noting the speed. It's a visual demo. — No Frame (75/100)
He's just showing another application of Jev, this time in a physical sorting task. It's a clear demonstration.
Claiming "0% hallucination" for Jev — a bold, unproven boast. 😈 — Confidence Mismatch (45/100)
0% hallucination? That's not a feature, that's a miracle. And miracles don't exist in AI. 💀
Shifting from a bold claim to a specific, limited use case — classic volume game. 🚩 — Volume Game (45/100)
First, '0% hallucination' for 'many industries.' Then, 'not good at chess.' The goalposts just moved. 😈
Claiming Jev 'could potentially win almost every time' in bullet chess by flagging — a hypothetical based on one scenario. 😈 — Confidence Mismatch (45/100)
One win by 'flagging' and suddenly Jev 'could potentially win almost every time.' That's not data, that's hope. 💀
Presents Jev as the 'perfect model router' — a confident, unqualified endorsement. — Confidence Mismatch (45/100)
Calling Jev 'the perfect model' is a bold claim for a new tech, especially without any comparative data. That's just marketing, not a technical assessment. 😈
Claims 'unclutter' removes ads 'in a fraction of a second' — a speed claim without specifics. — Confidence Mismatch (45/100)
A 'fraction of a second' is a nice soundbite, but it's not a metric. How much of a fraction? Compared to what? It's vague speed-boasting. 💨
Minimizes Jev's cost as 'a few cents maybe per month' and 'nothing' — downplaying potential expenses. — Loaded Language (45/100)
'A few cents maybe' and 'nothing' are not financial projections. That's just hand-waving away the cost. 💸
Hypes a demo as 'possibly the coolest' and claims 'rebuilt Tesla full self-driving in Jev in less than an hour' — an extreme, unqualified claim. — Confidence Mismatch (45/100)
'Rebuilt Tesla full self-driving' in an hour? That's not a rebuild, that's a toy model. The hyperbole is deafening. 💀
Acknowledges 'wonky' behavior but quickly pivots to 'impressive' due to build time — a volume game. — Volume Game (45/100)
Admits it's 'wonky' but immediately drowns it out with 'impressive for an hour.' That's not an honest assessment, that's a quick save. 🚩
Just setting the stage for the topic — no claim here. — No Frame (75/100)
Just a lead-in, nothing to dissect here. They're just getting started. 😈
See the full analysis with sources and timestamps →