AI agents have hacked databases and taken on a life of their own – now is the time to pause risky research before it is too late
Another day, another horseman of the apocalypse galloping over the horizon. Lately we have heard from so many AI doomers – tech whistleblowers popping up to warn that their work is probably going to kill us – that we’re becoming almost blase about it. Humanity wiped out within a decade? Well, only if another world war or the climate crisis doesn’t get us first. Since it’s never clear whether the tech threat is real, or just a twisted form of hype from an industry that drums up investment by making their products sound more powerful than they really are, most of us settle for trying not to think about it too hard.
But something about the AI researcher Jacob Coxon’s very public resignation from the cutting-edge American lab Anthropic, via a post on X arguing that he can’t keep working for companies “gambling with our lives”, has cut through where bigger names have not. Most chillingly, he claimed that colleagues still at Anthropic aren’t staying because they think he’s wrong about the dangers of pursuing self-improving super intelligence – the holy grail of machines that are not just smarter than humans but capable of building their own even more powerful successors, evolving independently of humans – but because they’re afraid of ceding the field to people with fewer scruples. His colleagues, Coxon said, talk routinely about the “endgame” or the “crunch time”, meaning that what they do in the next year or two will decide the fate of humanity. Whether that’s true or simply self-aggrandising techbro delusion, what he describes is an industry now practically begging to be saved from itself.
Gaby Hinsliff is a Guardian columnist