- Applied reviewer-requested changes - Quality gate pass (fix-from-feedback) Pentagon-Agent: Auto-Fix <HEADLESS>
35 lines
No EOL
2.7 KiB
Markdown
35 lines
No EOL
2.7 KiB
Markdown
---
|
|
type: claim
|
|
claim_id: no_frontier_ai_above_d_existential
|
|
title: No frontier AI company scores above D in existential safety despite active AGI development programs
|
|
description: FLI's Summer 2025 index shows all seven major frontier AI companies (Anthropic, OpenAI, Google DeepMind, Meta, xAI, DeepSeek, Mistral) received D or below in existential safety while actively pursuing AGI, demonstrating universal failure to implement adequate safeguards against catastrophic risk.
|
|
domains:
|
|
- ai-alignment
|
|
confidence: likely
|
|
created: 2026-03-10
|
|
tags:
|
|
- existential-risk
|
|
- frontier-ai
|
|
- ai-safety
|
|
- agi
|
|
---
|
|
|
|
The Future of Life Institute's Summer 2025 AI Safety Index evaluated seven major frontier AI companies—[[Anthropic]], [[OpenAI]], Google DeepMind, Meta, xAI, DeepSeek, and Mistral—and found that none scored above D in existential safety, despite all actively pursuing AGI development. This represents a universal failure across the industry to implement adequate safeguards against catastrophic risk.
|
|
|
|
The index assessed companies across multiple safety dimensions including dangerous capability testing, governance structures, and accountability mechanisms. The universal D-or-below rating in existential safety indicates systemic rather than company-specific failures, suggesting that competitive dynamics prevent even safety-focused organizations from prioritizing long-term risk mitigation.
|
|
|
|
This finding directly supports claims that [[voluntary safety pledges cannot survive competitive pressure when racing toward AGI]] and that [[no research group is building alignment through collective intelligence despite theoretical advantages]]—the index's evaluation framework omits collective intelligence approaches entirely, which is consistent with the field not recognizing CI as a viable alignment strategy.
|
|
|
|
## Evidence
|
|
|
|
- **FLI AI Safety Index Summer 2025**: All seven frontier companies rated D or below in existential safety
|
|
- **Companies evaluated**: Anthropic (D), OpenAI (D), Google DeepMind (D), Meta (F), xAI (D), DeepSeek (F), Mistral (F)
|
|
- **Evaluation framework**: Assessed dangerous capability testing, governance, accountability, and safety practices
|
|
- **Collective intelligence omission**: Index framework does not evaluate CI-based alignment approaches
|
|
|
|
## Cross-references
|
|
|
|
- [[anthropic scores C+ overall and D in existential safety making it the highest-rated frontier AI lab despite positioning as safety-first]]
|
|
- [[only three frontier AI companies conduct substantive dangerous capability testing despite universal claims of responsible development]]
|
|
- [[voluntary safety pledges cannot survive competitive pressure when racing toward AGI]]
|
|
- [[no research group is building alignment through collective intelligence despite theoretical advantages]] |