In a social media thread Tuesday night, Jacob Coxon announced his departure from Anthropic, warning that frontier AI companies are 'gambling with our lives' by developing systems that they 'earnestly believe... could kill us all by the end of the decade.' Coxon attributed the existential risk not to current models but to the impending prospect of 'self-improving superintelligence' creating 'superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources.'
He expressed concern that others working on these models have either not 'internalized the civilizational stakes' or believe they need to 'speedrun' the race to superintelligence to prevent an irresponsible party from getting there first.
Lending significant weight to Coxon's warning, Hubinger responded on social media to say that 'Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade.'
The episode stands out in a pattern where high-profile AI researchers often depart frontier labs to pursue startups or protest business models. Coxon's public warning, bolstered by a senior colleague's agreement, has intensified ongoing debates about AI safety and the pace of development at leading labs.