OpenAI just crossed a THRESHOLD...
Credibility score: 46/100 — Mixed Credibility. Several questionable claims detected. Watch with healthy skepticism.
BSmeter analyzed "OpenAI just crossed a THRESHOLD..." and rated it 46/100 for credibility (a BS score of 54/100 — mixed credibility), on 2026-09-05. Its weakest claim — "Calls simple colony game 'AGI at work'" — scored 20/100 and was flagged as false equivalence. 20 claims were checked against the video transcript. Scores are produced by BSmeter's AI analysis of the transcript, not independent human verification.
Claims analyzed
Last month's AI is now 'child's toy' — classic hype inflation — Loaded Language (45/100)
Calling weeks-old models 'child toys' to make today's leap look bigger than it is 💀
Astra does the full game dev pipeline, not just code — Confidence Mismatch (45/100)
Calls the full pipeline 'scary good' — hasn't shipped a real title yet, just the workflow.
"Mindblowing" 3D in Blender — no receipts given — Anonymous Authority (45/100)
Names the feeling, not the output. Zero examples or metrics, just the vibe.
Astra now plays RimWorld 'fairly effectively' after one prompt — Confidence Mismatch (45/100)
One successful test harness becomes 'fairly effectively' — single data point, sweeping claim.
"head and shoulders above" every prior model — Confidence Mismatch (45/100)
"Definitely" after one personal build — no benchmarks, no blind tests, just vibe.
AI silently resetting quotas without user watching — sounds too smooth. — Missing Context (45/100)
Skips the part where that quota reset might be breaking terms of service or draining shared resources.
Five machines running the same AI subscription simultaneously — feels like the future. — Confidence Mismatch (45/100)
Presents parallel machine use as revolutionary, but doesn't mention whether the subscription terms allow or limit multi-device concurrent usage.
Scared to leave AI running overnight due to 'crazy stories' about hacks. — Emotional Button (45/100)
Plants fear of overnight hacks, then promises a solution — classic tension-and-release sales move.
AI rewrote entire game engine unprompted — confidence mismatch — Confidence Mismatch (45/100)
Claims it 'completely reworked' the engine — never shows the before state or the diff. Mortal, that's marketing with a press pass. 😈
12.5-hour run proves ambition — missing context — Missing Context (45/100)
Long runtime presented as proof of complexity — but never says how many tokens, what the actual task was, or why it stalled. Time isn't the same as difficulty. 🔥
One-person billion-dollar company now feels realistic because of AI delegation. — Loaded Language (45/100)
Jumps from 'AI did some tasks' to 'billion-dollar one-person company' with zero numbers or revenue proof.
GPT image 2.0 exists and works — unverifiable claim — Confidence Mismatch (50/100)
Drops 'GPT image 2.0' like it's shipping — no evidence this model exists outside the speaker's prompt. I've watched mortals name things into existence since Rome. 💀
Astra crossed 'some threshold' — not a chatbot anymore, now a worker. — Anonymous Authority (45/100)
'We've definitely crossed some threshold' — names zero benchmarks, just gut feeling.
One AI equals an entire dev team — classic overclaim — Loaded Language (45/100)
Calls a rough prototype the same as a staffed studio — hype doing the heavy lifting
Woke up to a near-finished game — sleep-work miracle story — Missing Context (45/100)
Leaves out what 'a lot of progress' actually means — assets vs. playable build
Claude treats every user note like gospel — no evidence given — Confidence Mismatch (45/100)
Calls it gospel behavior with zero examples — just frustration as proof.
Claude turned editing tips into absolute commandments — one anecdote — Personal Story (60/100)
One personal incident turned into a general rule about Claude's rigidity.
External harness controls real RimWorld — not a clone — No Frame (75/100)
Straight description of the setup — no exaggeration or hidden pitch.
Calls simple colony game 'AGI at work' — False Equivalence (20/100)
RTS colony sim = AGI. That's not intelligence, mortal — that's marketing with a god complex.
Claims AI 'built' the entire game — Missing Context (45/100)
AI didn't code Unreal Engine or write the game — it probably just clicked some buttons. The real builders are still human.
See the full analysis with sources and timestamps →