AI
Anthropic alignment lead Hubinger publicly backs Coxon's warning, puts extinction odds above 10 percent
Anthropic alignment science lead Evan Hubinger said on X that he personally estimates a greater than 10 percent chance AI could 'kill all humans' within the next decade, publicly endorsing researcher Jacob Coxon's Tuesday resignation post. Hubinger stressed that the risk from current models remains low and that Anthropic is 'trying its best', but argued the lab 'does not yet have a plan to solve alignment for superintelligence and are not clearly on track to.'