Anthropic researcher Jacob Coxon is quitting and publicly warning that frontier AI companies are gambling with human survival by pursuing self-improving superintelligence. His concerns are echoed by Anthropic Alignment Science lead Evan Hubinger, who estimates over a 10% chance of AI causing human extinction within the next decade. The warnings align with Anthropic's own alignment team report acknowledging potential catastrophic risk from future more capable models.
Background
Anthropic is one of the leading frontier AI labs, known for its focus on AI alignment and safety research. The debate over AI existential risk has intensified as models become more capable, with some researchers advocating for pause or slowdown and others pushing for accelerated development.
- Source
- Ars Technica
- Published
- Sep 10, 2026 at 12:59 AM
- Score
- 7.0 / 10