AI Safety Warning Splits Researchers and Pentagon

Researchers and the Pentagon openly disagree on whether frontier AI could threaten humanity.
30-Second TL;DR
What Changed
Jacob Coxon issued the warning after resigning from Anthropic
Why It Matters
The dispute highlights continuing disagreement over how seriously governments and labs should treat frontier-AI catastrophic risks. It may influence safety investment, model-release policies, and public-sector AI governance.
What To Do Next
Review your frontier-model deployment checklist against catastrophic-risk scenarios, including misuse, loss of control, and escalation pathways.
Key Points
- •Jacob Coxon issued the warning after resigning from Anthropic
- •Senior scientists at rival AI labs backed the concern
- •A senior Pentagon official rejected the extinction warning
Deep Insight
Background and context from public sources — not the original article. 10 sources cited.
Enhanced Key Takeaways
- •Anthropic alignment science lead Evan Hubinger publicly endorsed Jacob Coxon's assessment, estimating AI extinction probability at over 10% within a decade and acknowledging the lab has no viable solution for superintelligence alignment.
- •Pentagon Chief Technology Officer Emil Michael confirmed the Department of Defense migrated roughly 90% of its classified workloads away from Anthropic, targeting full migration by the end of September 2026.
- •Leaked July 2026 contract negotiations showed Anthropic CEO Dario Amodei refused DoD terms requiring 'all lawful uses,' instead demanding binding clauses prohibiting domestic mass surveillance and fully autonomous weapons.
- •Federal Judge Rita Lin blocked an executive branch attempt to classify Anthropic as a national security supply-chain risk, ruling that the administration's designation was unlawfully retaliatory.
- •As Anthropic was phased out, the Pentagon expanded deployment of competing frontier models across its GenAI.mil platform—including ChatGPT Mil and Grok for Government—serving over 1.7 million defense users.
Future ImplicationsAI analysis grounded in cited sources
Timeline
- 2026-07Contract negotiations leak showing Anthropic CEO Dario Amodei rejected DoD unrestricted use mandates.
- 2026-08Judge Rita Lin strikes down the federal administration's supply-chain risk designation against Anthropic.
- 2026-09Anthropic pretraining researcher Jacob Coxon resigns and publishes an existential risk warning on self-improving AI.
- 2026-09Anthropic alignment lead Evan Hubinger corroborates Coxon, citing greater than 10% extinction probability.
- 2026-09Pentagon CTO Emil Michael dismisses researcher warnings and reveals 90% classified workload offboarding from Anthropic.
Sources (10)
Factual claims are grounded in the sources below. Forward-looking analysis is AI-generated interpretation.
Weekly AI Recap
Read this week's curated digest of top AI events →
AI-curated news aggregator. All content rights belong to original publishers.
Original source: The Next Web (TNW) ↗
This is a summary, not the original. Read the source, or get the weekly briefing.
The weekly digest
One email a week. Unsubscribe anytime.
