Researcher leaves OpenAI and Anthropic, warns of over 10% AI extinction risk
Jacob Coxon, who worked on pretraining at both OpenAI and Anthropic, has resigned and says the two labs are knowingly accepting a serious risk of human extinction from advanced AI. His former Anthropic colleague Evan Hubinger estimates the chance that a misaligned superintelligent system wipes out humanity within ten years at more than one in ten. The claims add to ongoing internal debate at leading labs about how much danger frontier models pose.