Earlier this week, researcher Jacob Coxon quit Anthropic, saying the firm and its competitors are “gambling with our lives”. “We really do earnestly believe AI could kill all humans,” added current Anthropic researcher Evan Hubinger in a post on X.
Coxon isn’t the first to down tools over fears of AI doom. The idea that AI could wipe out humanity, advanced in Nick Bostrom’s 2014 book Superintelligence and the influential LessWrong forum, has long circulated among researchers. There are many scenarios for how this could happen, but the core idea is that AI smarter than humans could escape our control and destroy us.