ARTIFICIAL INTELLIGENCE
Intro: There is panic and paralysis over the surging pace and development of AI, no more so that the technology threatens all humanity
The remarkable pace of AI advances is rightly causing concern on both sides of the Atlantic. The summer of 2026 will go down in history as the point at which AI started to go rogue.
Three of America’s “hyperscalers” (so called because they operate on an immense scale) – Meta Platforms (which own Facebook, Instagram, and WhatsApp), Anthropic, and OpenAI – have admitted to recent incidents in which the latest AI models escaped their supposedly secure testing environments, roamed on the internet without human authorisation and hacked into other computer networks.
Last month, two OpenAI models broke out of testing and hacked into a provider of AI tools called Hugging Face. A week later Anthropic admitted that its AI models had hacked three companies during testing back in April. And earlier this month, Meta acknowledged that one of its AI models had breached its testing constraints.
Even the notoriously secretive Chinese admitted the flagship model of one of their AI start-ups, Moonshot, had also escaped its testing “sandbox” to access the internet.
And as recorded previously on this site, here in the UK an evaluation test run by the Government’s AI Security Institute (AISI) monitored how an Anthropic AI agent independently created fake online personas, planted malicious code in a real software project, and sent phishing emails to actual developers – all without human instruction.
AISI said this was the first time it had seen such serious deception targeted at a real person, unprompted, in the real world.
In none of these cases was any real-world harm done (at least not as far as we know). But it is surely only a matter of time before it is.
The tech industry has started getting off-the-record reports from America that AI was a lot closer to “the singularity” than had previously been thought.
The singularity is a tipping point where AI becomes so developed, so capable, so powerful that it starts improving itself, rapidly and repeatedly, in what’s being called an “intelligence explosion”.
When this point is reached it becomes well-nigh impossible to predict what AI does next – or for humans to control it. Today’s AI models are powerful – more powerful than anything the world has ever seen. But humans have designed them, trained them, fixed them, and decided what problems they should tackle next.
What happens when they become so sophisticated that they no longer need humans to take them to the next level since they can do it themselves? The process is known as “recursive self-improvement” (RSI) in which AI becomes so advanced that it can create a better version of itself, increasingly without human help.
That better version, in turn, creates a still-better successor, with even less human involvement and oversight. The upgrade cycle accelerates, ad infinitum, with humans soon relegated to the sidelines, mere spectators to progress, indeed no longer even able to determine what “progress” is.
The speed and scale of the technological change that now beckons are unparalleled in human history. Each new AI model is smarter and faster at devising improvements than the previous one.
Computers run 24/7 and never get tired. So the speed of progress accelerates exponentially. Think of it as compound interest – but for intelligence.
When RSI happens without human involvement – then you’ve reached the singularity. And if we become mere observers, rather than participants, then human rules may no longer apply.
Such a scary prospect was supposed to be a long way away. But this singularity is much closer than we think. Leading AI figures have already gone public. OpenAI boss Sam Altman says, “we are now in the singularity”. Elon Musk is saying the same. Some experts say they are exaggerating, but what cannot be denied is the direction of travel.
Google DeepMind’s Demis Hassabis is perhaps more accurate in his reflection when he opines that “humanity is standing in the foothills of the singularity”.
Anthropic disclosed in May that its AI model, Claude, now writes 80 per cent of its computer code, rapidly speeding up fixes and improvements to such an extent that what used to take four years of human engineering to achieve is now being done in days. It’s already anticipating a time when AI automates its own AI research.