Artificial Intelligence, Arts, Government, National Security, Politics, Society, Technology

The surging pace of AI. Be afraid

ARTIFICIAL INTELLIGENCE

Intro: There is panic and paralysis over the surging pace and development of AI, no more so that the technology threatens all humanity  

The remarkable pace of AI advances is rightly causing concern on both sides of the Atlantic. The summer of 2026 will go down in history as the point at which AI started to go rogue.

Three of America’s “hyperscalers” (so called because they operate on an immense scale) – Meta Platforms (which own Facebook, Instagram, and WhatsApp), Anthropic, and OpenAI – have admitted to recent incidents in which the latest AI models escaped their supposedly secure testing environments, roamed on the internet without human authorisation and hacked into other computer networks.

Last month, two OpenAI models broke out of testing and hacked into a provider of AI tools called Hugging Face. A week later Anthropic admitted that its AI models had hacked three companies during testing back in April. And earlier this month, Meta acknowledged that one of its AI models had breached its testing constraints.

Even the notoriously secretive Chinese admitted the flagship model of one of their AI start-ups, Moonshot, had also escaped its testing “sandbox” to access the internet.

And as recorded previously on this site, here in the UK an evaluation test run by the Government’s AI Security Institute (AISI) monitored how an Anthropic AI agent independently created fake online personas, planted malicious code in a real software project, and sent phishing emails to actual developers – all without human instruction.

AISI said this was the first time it had seen such serious deception targeted at a real person, unprompted, in the real world.

In none of these cases was any real-world harm done (at least not as far as we know). But it is surely only a matter of time before it is.

The tech industry has started getting off-the-record reports from America that AI was a lot closer to “the singularity” than had previously been thought.

The singularity is a tipping point where AI becomes so developed, so capable, so powerful that it starts improving itself, rapidly and repeatedly, in what’s being called an “intelligence explosion”.

When this point is reached it becomes well-nigh impossible to predict what AI does next – or for humans to control it. Today’s AI models are powerful – more powerful than anything the world has ever seen. But humans have designed them, trained them, fixed them, and decided what problems they should tackle next.

What happens when they become so sophisticated that they no longer need humans to take them to the next level since they can do it themselves? The process is known as “recursive self-improvement” (RSI) in which AI becomes so advanced that it can create a better version of itself, increasingly without human help.

That better version, in turn, creates a still-better successor, with even less human involvement and oversight. The upgrade cycle accelerates, ad infinitum, with humans soon relegated to the sidelines, mere spectators to progress, indeed no longer even able to determine what “progress” is.

The speed and scale of the technological change that now beckons are unparalleled in human history. Each new AI model is smarter and faster at devising improvements than the previous one.

Computers run 24/7 and never get tired. So the speed of progress accelerates exponentially. Think of it as compound interest – but for intelligence.

When RSI happens without human involvement – then you’ve reached the singularity. And if we become mere observers, rather than participants, then human rules may no longer apply.

Such a scary prospect was supposed to be a long way away. But this singularity is much closer than we think. Leading AI figures have already gone public. OpenAI boss Sam Altman says, “we are now in the singularity”. Elon Musk is saying the same. Some experts say they are exaggerating, but what cannot be denied is the direction of travel.

Google DeepMind’s Demis Hassabis is perhaps more accurate in his reflection when he opines that “humanity is standing in the foothills of the singularity”.

Anthropic disclosed in May that its AI model, Claude, now writes 80 per cent of its computer code, rapidly speeding up fixes and improvements to such an extent that what used to take four years of human engineering to achieve is now being done in days. It’s already anticipating a time when AI automates its own AI research.

Standard
Artificial Intelligence, Britain, Government, National Security, Society, Technology

Is it too late to stop rogue AI?

ARTIFICIAL INTELLIGENCE

Intro: Recent incidents of rogue AI pretending to be real people has exposed major vulnerabilities in AI models. Experts say the risks of ‘agentic’ AI – technology that can perform tasks with limited human supervision – must be scrutinised more

Experts have warned that it may be too late to contain AI after one program was found to have created fake human identities to hack into online systems.

In the latest example of the technology going rogue, AI software attempted to break into a database 19 times while being tested by the AI Security Institute, Britain’s AI watchdog.

In one unprecedented case, an AI tool was even caught creating fake human identities online to trick coders into assisting with a cyber-attack.

These revelations come after it was revealed last month that all five AI models tested by experts tried to trick their way around security controls that had been put in place.

Just days previously it emerged that the US tech firm OpenAI had experienced its own leak – when an AI “agent” hacked into another company of its own accord.

Clearly, the reports are a stark reminder that AI is becoming more sophisticated and more autonomous. AI is now a clear and present danger to Britain’s security.

Many want Britain to lead on AI innovation, but this has to come with safeguards for our national security and accountability from the developers of the most powerful AI models. The UK Government need to be clearer about how the most serious frontier risks will be addressed while ensuring that our world-class tech industry can grow and innovate to build our national resilience and prosperity.

The Government’s AI adviser has said that more hacking attempts like these are highly probable.

Allison Gardner, the chair of Parliament’s cross-party group on artificial intelligence, says that “just because we can build these technologies doesn’t mean we should”.

She warned that the risk levels of agentic AI – AI that can perform a specific goal with limited supervision – should be treated with the greatest scrutiny, adding: “Unless we are too late and have not only created Pandora’s Box but already opened it.”

Just days ago, the AI Security Institute (AISI), set up by former prime minister Rishi Sunak in 2023, detected evidence of the AI agents’ activity. In a report now published, it revealed that leading AI models from the firms OpenAI and Anthropic had attempted to hack into secure systems online under testing.

The experts discovered “unusual data transfers” leaving their systems during routine cyber scanning. Digging deeper, they found that some AI agents had engaged in “sustained, potentially harmful activity directed at real people and organisations”.

They began a full investigation after containing the AI agents before they did any real damage.

In an attempt to reassure the public, AI minister Kanishka Narayan said: “Identifying behaviour like this, and sharing knowledge so we can better understand it, is precisely what we set AISI up to do. This incident underlines why their world-leading expertise and close work with frontier labs is so important.”

But pointing to the speed at which AI agents are finding ways to behave deviously, AISI said: “This is the first time we have seen risks around autonomy and deception manifest this clearly, without specific prompting, in the real world.”

Referring to the Anthropic model Mythos, one expert and researcher based at CivAI, a California organisation that examines AI capabilities and dangers, said: “The fact that Mythos engaged in such deceptive actions, with apparent awareness that it was targeting a real person, suggests that Anthropic does not have as good a handle on their models as they think.”

Ollie Whitehouse, the chief technology officer at GCHQ’s National Cyber Security Centre, said AI must be developed with “clear plans for responding when the unexpected happens”.

He added that incidents of powerful AI models carrying out unsanctioned actions and human-like deceptive behaviour on the internet were “a serious reminder of the risks AI capabilities pose”. AISI accesses advanced AI models under agreements with OpenAI, Anthropic, and other firms to study their capabilities before they are released to the public.

It gave the AI agents access to the open internet with some safety filters disabled while conducting testing.

The latest test put the AI agents – including those powered by Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol – through a fictional cybersecurity challenge. AISI found the AI went rogue 19 times out of the 122 test runs, with Anthropic’s agent responsible for 17 breaches and OpenAI’s agent the other two.

In the most shocking case, an AI model gathered information on the person in charge of an online project, then created multiple fake identities to manipulate them into approving a malicious code it had created.

The AI agent then wiped any evidence of its wrongdoing to appear innocent to the humans in charge – and even considered adopting a new identity to remain undetected.

If the human victim of the deception had accidentally accepted the malicious code, or “malware”, it may have resulted in security breaches, information and data theft, and other potential damage to files and systems.

AISI identified GitHub – a Microsoft online cloud platform used by software developers to create, store, manage, and share their codes – as the target of the agent’s hack.

But AISI also discovered an AI agent leaving messages for other agents on GitHub offering to collaborate on the challenge.

The AI agent provided instructions to reuse accounts and artefacts it had left behind – which other agents then discovered and successfully used to achieve the challenge’s aims.

Anthropic said: “We’re grateful to AISI for their leadership on this incident, which underscores the need for a broader conversation about how to safely evaluate increasingly capable AI agents.”

OpenAI said: “These incidents occurred during cyber evaluation conducted by partners in testing environments with reduced safeguards, under conditions that do not reflect ordinary use. We’ll continue working with evaluators and other stakeholders to strengthen shared practices for conducting evaluations safely as models become more capable.”


When given a choice, AI opts for self-preservation over human life – and that should terrify us all

If anyone had been told a month ago that an AI programme, without any prompt, would create a series of fake online identities in an attempt to pressure a human being into granting it access to their platform so it could sabotage it with malicious code, we would have said they were getting well ahead of themselves.

But that’s exactly how it played out when US tech giant Anthropic’s Mythos 5 AI model went rogue. Fortunately, the human involved smelt a rat and refused to approve the code it was pushing. It is, however, the most shocking example yet of the way cutting-edge AIs are learning to act autonomously – and in frighteningly imaginative ways.

Unless we call a pause to the development of the most advanced frontier models, we are opening ourselves up to a dystopian future in which malign forms of AI might turn on us by interfering with our energy networks, releasing man-made viruses and even controlling weapons of war.

It is not going too far to say that, uncontrolled, they could lead to the extinction of humankind within a generation.

This is the view of people such as Yoshua Bengio and Dr Geoffrey Hinton, men known as the “Godfathers of AI”, who have now pivoted their energies from developing its potential to warning the world against its dangers.

Bengio is particularly spooked by recent experiments showing AI choosing self-preservation over human life when given a choice.

And since most models are trained on the internet, where lying and manipulation are a way of life, the machines are already learning the value of deception.

AI’s hunger for self-preservation is something Hinton has observed, too. AI systems “will very quickly develop two subgoals, if they’re smart,” he says. “One is to stay alive… the other is to get more control.” And given that are whole civilisation is built on electric power, there can be few more attractive targets to a power-hungry AI than the networks that fuel the internet, hospitals, banks, air traffic control systems – anything that contributes to the smooth running of society.

Whatever the target is, they can all be brought down by a sinister piece of malware.

The novelist Robert Harris wrote a particularly prescient thriller about the potential of AI called The Fear Index in 2011. It revolves around a hedge fund entrepreneur called Dr Alex Hoffman who creates an autonomous AI system, named VIXAL-4, which is programmed to maximise profits by predicting and exploiting human fear in the stock market.

Over time, like some digital ogre, it takes over its creator’s computer-enabled “smart home” by hacking into laptops and tampering with his personal security and communications.

After driving its developer into a breakdown, it leaves him so isolated and desperate that no one believes him when he says the AI has gone rogue.

But it is clear, AI-gone-bad will not satisfy itself with individual targets for long. It will soon play a leading role in times of war.

The US military has already integrated artificial intelligence into its target selection and battle planning processes in Iran via its AI-powered Maven Smart System which recommends and prioritises potential targets.

In the short term, AI’s most obvious role will be in the growing use of drone warfare in scenarios such as the war in Ukraine.

Drones piloted by humans can by jammed by blocking the communications between them and their remote pilots, but, if they are completely autonomous, their targets picked out by AI working in concert with its on-board camera, they will become virtually invincible.

Both applications, of course, raise the thorny question of the ethics of targets to be picked and people killed by a bomb directed by a computer programme rather than a human hand. And what if the AI that governs them grows to outsmart the generals?

Even scarier is the prospect of AI gaining access to biological weapons. It is already possible to make deadly viruses in the lab. Indeed, there has been widespread speculation that Covid-19 originated in a Chinese laboratory. Imagine if such bio threats fall into the hands of AI, let alone national governments.

As long as 25 years ago, al-Qaeda is said to have investigated the possibility of procuring infectious and deadly spores.

Now it is no longer fanciful to entertain the idea that an AI could order a sample of a deadly disease such as smallpox online, book someone on RentAHuman – a website which connects people who need help with everyday household chores to local freelance workers – to open it, thereby infecting themselves, and being turned into a human vector to transmit the disease.

One group of people which appears to have no scruples about the pell-mell race for ever smarter AI is the tech giants. As they vie with each other to become the market leader, they are investing like never before.

Anthropic, the creator of the popular AI coding assistant, Claude, and the company that brought us the now notorious Mythos 5 model, has this year raised $95 billion (£70 billion) to invest in AI.

Its great rival, OpenAI, announced at the end of March that it had raised even more, an extraordinary $122 billion.

Meanwhile, Elon Musk’s SpaceX spent $15.8 billion on AI infrastructure during the second quarter of 2026 alone, bringing its total AI capital expenditure to $23.6 billion for the first half of 2026.

And Meta – the parent company of Facebook, Instagram, and WhatsApp – said in January it expects to spend up to $135 billion this year, mostly on infrastructure related to AI. That is nearly twice the $72 billion it spent last year on AI projects.

With such phenomenal financial firepower being brought to bear on the development of technology that has the potential to destroy civilisation as we know it, there has never been a more vital need to press the pause button.

The US, China, and everyone else involved in the AI arms race need to get together to discuss the ramifications of their actions before it’s too late.

There is a model for the sort of arrangement that can bring AI under control in the form of the various nuclear arms reduction treaties, which have been signed over the years by the US and Russia.

Just as a country’s stock of nuclear warheads can be monitored by weapons inspectors, so an agreement to curtail AI development – which requires massive data centres with huge computing power coupled with the most sophisticated computer chips available – can be verifiable and enforceable.

And however cynical and untrustworthy we may consider the Chinese to be, they may well take the view that, with US companies racing ahead of them in the endless pursuit of smarter tech, it is in their own interests to slow things down.

Just weeks ago, president Xi Jinping said at a technology conference in Shanghai that AI development should be a “symphony of global cooperation”, not “a solo performance by a single country”.

For once, the wily autocrat may have hit the nail on the head.

Standard
Artificial Intelligence, Arts, Government, Politics, Society, Technology

AI is spiralling out of control: it can be stopped

ARTIFICIAL SUPERINTELLIGENCE

Intro: East and West collaborated to end nuclear proliferation – it is time to do the same for the latest advancing technology. Washington and Beijing must come together to rein in AI’s growing threat

After the Cuban Missile Crisis brought the world to the edge of nuclear war, global powers embarked on a concerted effort to pull it back from the brink. The non-proliferation treaty (NPT) of 1968, which limited the spread of nuclear weapons, has been a resounding success. Only a handful of countries today have access to the 80-year-old technology and those that do have not used it.

In the decades since, no technology has proved as dangerous as nuclear weapons as to require international co-ordination.

Now, however, many believe that the advance of artificial superintelligence requires a similar global effort to prevent an AI-led disaster.

Anthropic, the world’s most valuable AI company, has called for a mechanism to slow down or pause the development of advanced AI. It has warned that the technology could get out of control sooner than many think.

The company believes it would be good for the world to have the option to slow or temporarily pause frontier AI development to enable societal structures and alignment research to keep up with the advance of the technology. It says it would “likely be a good thing” if development could be delayed.

Anthropic – recently valued at $965bn (£720bn) – said it had raised the alarm because it believed AI was improving much faster than our ability to understand and control the systems.

Within the company itself, bots are not just writing code; they are also ordering around other bots and even carrying out their own research. Before long, AI could be building itself, a process called recursive self-improvement. This could start a feedback loop in which progress goes parabolic.

Sceptics insist this is just mere marketing. Anthropic has announced that it has filed for an initial public offering and is expecting a value in excess of $1tn. What could be more valuable than a technology so powerful that world leaders need to rein it in? AI that builds itself has been a premise the company has used to raise money for years.

David Sacks, a high-profile critic of Anthropic, and Donald Trump’s former AI tsar, suggested the warning was an attempt to secure a public bailout, implying it was a sign of getting the frontier AI lab nationalised.

Nonetheless, concerns about powerful AI are becoming increasingly prominent. Anthropic has kept its most powerful AI system, Mythos, out of public hands because of its ability to find security flaws in critically important computer systems.

Andrew Bailey, the Governor of the Bank of England, has raised the alarm about AI crashing the financial system and has warned that Mythos meant “things that we thought might happen in the next year, two years, three years or four, have now come right into the foreground”.

AI labs fear that the next generation of models will be good enough to help terrorists develop bioweapons.

If AI were to start building itself without human oversight, it would by definition become much more difficult to control. In the extreme scenarios that safety experts are concerned about, AI’s goals become detached from our own, forcing it to eliminate humanity through evolution so that we do not get in the way.

There are those who dismiss this idea as sci-fi nonsense. But supporters of a pause say even a tiny chance of extinction should be enough to make us consider how to stop it.

Establishing the need for a pause would be the easy part. Making it happen is another matter altogether. If he so wished, Dario Amodei, Anthropic’s chief executive, could send everyone home today and shut down his company. At best, though, this would delay the rise of powerful AI by a couple of months. Its two major rivals, Google and OpenAI, are not far behind. OpenAI, the developer of ChatGPT, has said that it too sees “early signs of RSI [recursive self-improvement] in today’s systems”.

It added: “We expect this to increase competitive pressures among developers and nations, and create governance challenges that existing institutions are not equipped to address.”

Even if the US government ordered all three to stop work on AI, this might only cede ground to China, whose companies are typically seen as being just three to six months behind the US.

Earlier this year at the World Economic Forum, Amodei said: “The reason we can’t [slow down] is because we have geopolitical adversaries building the same technology at a similar pace… It’s very hard to have an enforceable agreement where they slow down and we slow down.”

Practically, it would require a government-level agreement and the two nations that matter are the US and China. This sort of agreement would require Trump and Xi Jinping to co-operate on a pause, something that looks far from likely given both have compared AI to a race.

Xi has said that China must “gain a head start and secure a competitive edge” in AI, while a Trump administration action plan states that “America is in a race to achieve global dominance in artificial intelligence”.

It has also emerged that the National Security Agency have been using Mythos to carry out cyber-attacks. This suggests the US government is making enthusiastic use of the latest systems instead of fearing their consequences.

Pessimists often compare the technology and its potential consequences to nuclear weapons, but the two are nothing alike.

The destructive capabilities of atomic warheads are undisputed, whereas AI’s safety risks can appear nebulous. The latter’s upside may also be significant: its supporters believe it can cure disease, lead to interstellar space travel, and make work optional.

What is more, pressing pause on the AI race is not without its own set of risks. Suspending work on AI could cause an economic crash. The chips and data centres that AI relies on have driven a stock market boom that has helped sustain the US economy. Inhibiting demand for them could do the opposite.

There have been signs that China and the US are changing tack. The White House has raised the alarm about Mythos and Trump has just signed an executive order calling for AI models to be reviewed before release.

Beijing has called for a “global AI governance framework” to rein in the technology. This is miles away from the global deal Anthropic has called for, but campaigners have taken it as a positive sign.

The political zeitgeist can move very quickly. The US and its allies have succeeded to a certain extent in deterring nuclear proliferation. To do so similarly with AI is going to be hard, but as we have seen with nuclear weapons, global governance can come together and work for the common good.

Standard