The trick: Self-Marked
The new best open model beat GPT and Claude in every test.
Try finding one you can verify.
Meet Qwen3.8-Max, our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released
Before you read on. Your call?
TRUE, BUT
True, but
There’s more to this story.
Membership opens the full investigation, the strongest counterargument and what to do with what you’ve learned.
Start your free month →First membership: 30 days free, then A$89 a year. One introductory trial per customer. Card required; renews annually until cancelled. Cancel before the trial ends to avoid the first charge. Already a member? Sign in
Couldn't check your access. That's on us.
The trick has a name
We call it Self-Marked: graded by the party that benefits from the grade. You'll see it again. Learn to spot it →
Receipts
- Supports latent.space:
Vals Index: #2 among open models, 66.1 score (matched Claude Opus 4.7)
- Refutes yottalabs.ai:
every stronger claim than that currently traces back to Alibaba's own press materials
- Context glbgpt.com:
the table mixes official model reports, system cards, public leaderboards, and Qwen's in-house evaluations
Open the Receipts Pack → What each source proves, every figure traced, and what would change our verdict.