Seven independent experts graded every frontier AI lab on safety.
The best grade was a C+. Three labs got an F. The best existential-safety grade was a D+, and nobody reached a C. And the labs that used to promise to stop at red lines? They rewrote the red lines.
Every frontier AI lab scores C+ or below on the FLI AI Safety Index. No company passes on existential safety.
Before you read on. Your call?
VERIFIED
Report card
seven experts from UC Berkeley, Oxford, U Montreal, and others evaluated nine labs on 37 indicators across six domains through June 2026. The panel found Anthropic, OpenAI, Google DeepMind, and Meta weakened or voided pledges to pause if redlines are approached, replacing them with competitor-contingent conditions. Labs that once banned military applications now seek defense contracts. David Krueger called the lack of progress scandalous. Four of nine companies did not respond to the survey.
There’s more to this story.
Membership opens the full investigation, the strongest counterargument and what to do with what you’ve learned.
Start your free month →First membership: 30 days free, then A$89 a year. One introductory trial per customer. Card required; renews annually until cancelled. Cancel before the trial ends to avoid the first charge. Already a member? Sign in
Couldn't check your access. That's on us.
Receipts
- Supports futureoflife.org:
No company exceeds C-; most score D or below.
- Supports futureoflife.org:
Anthropic, OpenAI, Google DeepMind, and Meta have weakened or voided pledges to pause unilaterally.
- Supports futureoflife.org:
Companies previously banned military applications gradually reversed course.
- Context digitalapplied.com:
A C+ does not mean Anthropic's systems are safe to use in all contexts.
- Supports en.sedaily.com:
Companies are backing away from the red line commitments they previously made.
Open the Receipts Pack → What each source proves, every figure traced, and what would change our verdict.