Subscribe

The trick: Self-Marked

The company that grades AI for OpenAI, Anthropic, Google, Meta, and xAI just raised $40M at a $400M valuation.

It disclosed a customer relationship with the labs it evaluates.

Issue 1016 August 20264 receipts3 min

Vals AI raised $40M led by a16z at a $400M valuation on the pitch that it is the independent evaluator of AI. Its benchmarks appear in model cards from OpenAI, Anthropic, Google, Meta, and xAI.

Before you read on. Your call?

Artificial Lawyer reported Vals AI disclosed a customer relationship with one or more of the participants it evaluates. The benchmarks that appear in model cards are produced by a company billing the labs whose products carry those scores.

The twist

independence is the product. The customers are the subjects. The conflict is structural, not hidden. Vals disclosed it. But disclosure does not eliminate the conflict; it just puts it on the record.

40million USD Series A per Crypto Briefing
52% Finance Agent v2 accuracy for GPT-5.5

There’s more to this story.

Membership opens the full investigation, the strongest counterargument and what to do with what you’ve learned.

Start your free month →

First membership: 30 days free, then A$89 a year. One introductory trial per customer. Card required; renews annually until cancelled. Cancel before the trial ends to avoid the first charge. Already a member? Sign in

The trick has a name

We call it Self-Marked: graded by the party that benefits from the grade. You'll see it again. Learn to spot it →

Say this in tomorrow's meeting“The AI benchmark referee just raised $400M. Its headline stat is three months stale, and it disclosed customer relationships with the labs it grades. That is not independence. That is consulting.”

Receipts

  1. Supports vals.ai: We started Vals AI to solve this problem as the independent evaluator of artificial intelligence.
  2. Refutes artificiallawyer.com: customer relationship with one or more of the participants
  3. Supports cryptobriefing.com: $40 million in a Series A
  4. Context kucoin.com: Vals AI Finance Agent v2 Reveals GPT-5.5 Hits Just 52% Accuracy

Open the Receipts Pack → What each source proves, every figure traced, and what would change our verdict.

This story is a stable, citable object. If you can falsify a verdict,tell us. Corrections are loud here.