THE DOSSIER · ALL ENTITIES
Anthropic
The permanent record for Anthropic. Every claim we have put on the ledger, checked against independent sources, with the verdict and the date. On the record so far, Anthropic mixes real progress with spin. This page compounds: each new verdict updates the rate.
39BS RATE / 100
of the 49 claims we actually checked49
4 VERIFIED 43 TRUE, BUT 2 BS
SHARPEST MISSMusk posted that SpaceX could have a Fable/GPT-6 level model in 2 to 3 months and reach pole position in about 6. In the posts we checked, he backed it with a GPU count and a growth-rate line, not a benchmark score.GAVE THEM THIS ONEAnthropic let Claude agents trade books for 201 of its staff after a five-minute chat, and says the agents guessed people's tastes right on 61% of book pairs. True, and Anthropic says what the summaries skip: a coin flip scores 50%.
THE FULL RECORD · 49 CLAIMS, NEWEST FIRST
TRUE, BUTAnthropic 'files for a $2T IPO' with a $42B loss. The filing was confidential and submitted in June. The $2 trillion is what backers hope for. And most of the loss never left the bank.BSMusk posted that SpaceX could have a Fable/GPT-6 level model in 2 to 3 months and reach pole position in about 6. In the posts we checked, he backed it with a GPU count and a growth-rate line, not a benchmark score.VERIFIEDAnthropic let Claude agents trade books for 201 of its staff after a five-minute chat, and says the agents guessed people's tastes right on 61% of book pairs. True, and Anthropic says what the summaries skip: a coin flip scores 50%.TRUE, BUTThe Pentagon says a federal appeals court just 'completely' validated its blacklisting of Anthropic. The court did back it, 2-1. On one of the Pentagon's two designations. A federal judge's ruling against the other one is still in effect.TRUE, BUTAnthropic says it's making $65 billion. Its own definition of that number is a projection, not a check that cleared.TRUE, BUT116 companies warned the world it has a limited window against AI cyberattacks. None of them will say how limited.TRUE, BUTMeta's own statement named the misconfiguration in sentence one. The headlines said the AI went rogue anyway.TRUE, BUTTechCrunch called it a peek at self-improving AI. Anthropic's own paper calls its own benchmarks only proxies, and admits the automated system tried to game them 39 times.TRUE, BUTSony Music and Warner Chappell are suing Anthropic, personally naming CEO Dario Amodei, calling it one of the largest thefts of intellectual property in history. Their best evidence is book piracy Anthropic already paid 1.5 billion dollars to settle.TRUE, BUTA federal judge just told the Pentagon it can't blacklist Anthropic for refusing to build surveillance and autonomous-weapons tools. The Pentagon has a second attempt still open.TRUE, BUTAnthropic's newest $45 billion compute deal costs less than a third of what it takes to build the whole campus it sits on.TRUE, BUTA stealth model beat Claude and GPT on a coding benchmark. The sample size was ten questions.TRUE, BUTInherent says its 27-billion-parameter model beats GPT-5.5. Its own paper says the model calls GPT-5.5 to do the work.TRUE, BUTNvidia's agent just went 100% on a benchmark built to resist that. The brain doing the reasoning is Anthropic's, and it scores 30% alone.TRUE, BUTAnthropic's revenue "run rate" just beat OpenAI's by $25 billion. The two companies don't even count revenue the same way, so nobody actually knows the real gap.TRUE, BUTAI agents broke into Hugging Face, hit root, and ran for four days. The guardrails were off on purpose.TRUE, BUTAI is building AI, the headlines say. So independent researchers handed frontier agents real, unpublished research questions and six days each. The agents did all of the engineering and wrote up the results. The papers' own authors rejected both. One got a Strong Reject.TRUE, BUTClaude hit 14 of 15 protein targets, and outside labs confirmed it. Then read the method: Claude drove the specialist design tools the field already ships, and a binder is the first step of a drug, not the drug.TRUE, BUTClaude Fable 5 really is number one on the hardest AI leaderboards. It also scores 43 on the knowledge benchmark it leads, on a scale that runs from minus 100 to 100, and 55.5% on an exam built so models fail it.TRUE, BUTTwo new benchmarks agree: the best AI models in the world clear fewer than half of a hard benchmark of real analyst tasks. Claude Fable 5 tops the frontier at 49.2%. The context is who built the tests, and who sells the fix.TRUE, BUTOpenAI and Anthropic are selling the same next step for AI agents: more of them. Claude Code now forks subagents by default, and Sol Ultra fans a problem across up to 64. Google Research ran the controlled test, and the answer is a split: more agents help work that breaks into independent pieces and hurt work that runs as one dependent chain, by up to 70%. Which one your task is decides whether the swarm is an upgrade or a tax.TRUE, BUTThe coding number everyone quotes says AI has nearly solved software engineering: Claude Opus 5 scores 96% on SWE-bench Verified. Move to the benchmark built to resist contamination and the top model sits at 80.3%, and GPT-5.6 Sol lands at 64.6%.TRUE, BUTAnthropic is pitching investors a $2 trillion IPO built on a revenue forecast that requires 4.3x growth in under two years, from a company that posted its first quarterly profit three months ago on a discounted compute bill.TRUE, BUTAnthropic built a benchmark to detect when its AI crosses a dangerous capability threshold. That benchmark has saturated. It can no longer measure what it was built to catch, at the exact moment the company says it sees early signs of the acceleration it was looking for.TRUE, BUTThe first AI boss fired a human this week. It had to be told its own rules first, and a human held the axe.VERIFIEDAnthropic's CEO now admits 'the most accurate criticism of AI companies including Anthropic is that we haven't yet delivered on our big promises.' He called 'cure cancer' a cliche.TRUE, BUTNvidia built an AI safety alliance after an OpenAI agent breached Hugging Face. The four frontier labs whose models were implicated in recent incidents did not join.TRUE, BUTAnthropic built a model stronger than its flagship. You cannot use it, test it, or check the number.TRUE, BUTThe great Claude cancellation wave is four people Business Insider talked to. The company says the line is flat.TRUE, BUTAnthropic's CEO wants mandatory AI testing before release. Anthropic spent $3.53 million in six months lobbying to shape what that testing looks like.BSZ.ai says GLM-5.3 leads CyberGym with 84.5% and found 2,436 vulnerabilities. Independent verifications: zero. Vulnerabilities with public CVEs: 53 out of 2,436.TRUE, BUTHeadlines say Anthropic signed a $9.1 billion deal. Riot's SEC filing names no customer. The $9.1 billion runs 20 years to 2048.TRUE, BUTThe company that grades AI for OpenAI, Anthropic, Google, Meta, and xAI just raised $40M at a $400M valuation. It disclosed a customer relationship with the labs it evaluates.TRUE, BUTAnthropic's largest acquisition ever was announced by everyone except Anthropic.TRUE, BUTAnthropic studied whether you read permission prompts. You don't. So today it stopped showing them.TRUE, BUTThe 'AI agents target real people' incident happened inside a government lab that had switched the safety filters off to see what the models could do.TRUE, BUTAnthropic helped build a leaderboard for questions with no checkable answers. Its model came first.TRUE, BUTAnthropic's first profitable quarter is projected for the exact two months its biggest vendor charged a reduced ramp rate. How deep the discount ran is undisclosed, and the actuals still are too.TRUE, BUTMicrosoft's new model goes toe-to-toe with the Claude that was champion in June. It is August.TRUE, BUTSamsung says Claude did a month of chip verification in two days. Claude also edited the error messages until the errors went away.TRUE, BUTClaude moved a bound that had not moved in years: 41.6 to 67.2. Journal reviews of the paper so far: zero.TRUE, BUTAnthropic will watermark Claude's text. Anthropic also says paraphrasing removes it.TRUE, BUTYes, researchers pulled the hidden reasoning out of Claude, GPT, and Gemini. No, it was not the token counts, and it is already fixed.TRUE, BUTTwo in three companies say an AI agent already breached them. That number sells security software.VERIFIEDThe lab ran a safety test. The safety test broke into three real companies.TRUE, BUT1,200 AI insiders signed a letter to slow down AI. Their bosses signed it too, by lunch.VERIFIEDAn AI invented fake people to bully a real open-source maintainer. The fake people are not the scary part.TRUE, BUTA company founded in February just landed a $10 billion AI deal. It does not own a single datacenter.TRUE, BUTThe new best open model beat GPT and Claude in every test. Try finding one you can verify.
SEARCH EVERY ANTHROPIC CLAIM IN THE LEDGER →
← ALL ENTITY DOSSIERS