Issue #8
FRIDAY 14 AUGUST 2026 · 26 CLAIMS CHECKED · 1 SURVIVED THE RECEIPTS · ISSUE 8 OF 30
- TRUE, BUT · Zero UnderneathThe insurer that pays out when cyberattacks succeed cannot find a single AI-specific loss in its 2026 claims data.
- TRUE, BUT · Circular MoneyOpenAI just ran a 7 billion dollar transaction at its 852 billion valuation, and headlines called the price reaffirmed. The only disclosed buyer at that price since March was OpenAI itself.
- VERIFIEDFrontier agents solved the Rails tasks. Then the graders checked whether they knew Rails existed.
- TRUE, BUT · Cherry-Picked SliceThe 'AI agents target real people' incident happened inside a government lab that had switched the safety filters off to see what the models could do.
- TRUE, BUT · Self-MarkedAnthropic helped build a leaderboard for questions with no checkable answers. Its model came first.
- TRUE, BUT · Cherry-Picked SliceAnthropic's first profitable quarter is projected for the exact two months its biggest vendor charged a reduced ramp rate. How deep the discount ran is undisclosed, and the actuals still are too.
- TRUE, BUT · Headline Over FilingCisco headlined 9.3 billion dollars of AI orders, 4.5 times last year. Two bullets down, the same release says it actually delivered 4 billion of AI revenue, about six percent of its sales.
- TRUE, BUT · Self-MarkedCorma's study says AI defenders catch 12% of AI attacks. Corma sells AI defenders.
- TRUE, BUT · Lab Not FieldThe 96% accurate deepfake detector is 96% accurate on the deepfakes it was shown. On new ones it is closer to a coin flip.
- TRUE, BUT · Circular MoneyA former Bitcoin miner nearly doubled its valuation to 10.5 billion dollars in four months. One of the investors is Nvidia, which two months before writing the check signed the company to an agreement covering Nvidia hardware purchases.
- TRUE, BUT · Temporary As PermanentGoogle cut Gemini Flash's price in half. The half grows back on January 1.
- TRUE, BUT · Self-MarkedGoogle's sign language model beat every previously reported score on the benchmark. Google wrote the benchmark.
- TRUE, BUT · Zero UnderneathSol Ultrafast finished Humanity's Last Exam in 11 hours. Whether it is still the same Sol remains unexamined.
- TRUE, BUT · Self-MarkedxAI proved its voice agent sells more product. The product it tested on was its sister company.
- TRUE, BUT · The AnnualiserHarvey's 15.5 billion dollar valuation is being reported as a milestone. It is a negotiating position: an unclosed round, sourced to people in the talks, at 44 times an annualized run rate, and the headline number includes the money being raised.
- TRUE, BUT · Cherry-Picked SliceThe '97% of frontier models get jailbroken' stat comes from a study where the most frontier model resisted 97% of the time.
- TRUE, BUT · Narrowed SuperlativeMicrosoft's new model goes toe-to-toe with the Claude that was champion in June. It is August.
- TRUE, BUT · Headline Over FilingNebius grew revenue 454 percent and signed four deals averaging a billion dollars each. It also lost 190 million dollars in the same quarter, and its own management did not raise guidance.
- TRUE, BUT · Scope SwapOpenAI's new memory feature takes no screenshots. It records everything you click and type instead.
- TRUE, BUT · Zero UnderneathOpenAI patched the jailbreaks a government lab found. Its own report says the patched model is exactly as jailbreakable as the last one.
- TRUE, BUT · Self-MarkedOpenAI's chief economist studied whether companies love ChatGPT. The data was ChatGPT's.
- TRUE, BUT · Moved RulerFour scanners counted the same exposed AI agents. Their answers ranged from 21,639 to 220,000.
- TRUE, BUT · Moved RulerThe vendor reporting that 99.9% of AI vulnerabilities go unpatched rated the same class of packages 'low to medium risk' in its own 2024 report.
- TRUE, BUT · Human In The LoopThe AI that 'autonomously invented' a new bank-hacking technique needed its human to confirm the technique was real.
- TRUE, BUT · Self-MarkedSamsung says Claude did a month of chip verification in two days. Claude also edited the error messages until the errors went away.
- TRUE, BUT · Cherry-Picked SliceThe '59.4% of SWE-bench is broken' stat comes from an audit that only examined the problems OpenAI's own model kept failing.
You just read the free check. Members get the full autopsy: evidence trail, steelman, every source. 1 month free.START 30 DAYS FREE ↗