Anthropic Alignment Lead: Over 10% Chance AI Kills Us All
Evan Hubinger, who leads alignment stress-testing at Anthropic, says he personally puts the odds of AI killing all humans at over 10% within the next decade — and that the company "does not yet have a plan to solve alignment for superintelligence". He was backing Jacob Coxon, a pretraining researcher who quit on September 8 accusing Anthropic and OpenAI of "gambling with our lives".
TIME / TechCrunch
Read more →