Artificial Intelligence, Science, Society, Technology

Tech bosses have hidden motives in slowing AI progress

ARTIFICIAL INTELLIGENCE

Intro: What’s the real reason tech bosses are warning us about their AI creations? It’s certainly nothing to do with believing we are on the verge of wiping out humanity

The terrifying message from Silicon Valley is that we’re all doomed as the developers of artificial intelligence foretell the apocalypse for which they will themselves be responsible.

A high-profile British employee of Anthropic announced he was leaving over fears his company and its main competitor, OpenAI, were “gambling with our lives” by rushing to create a digital “superintelligence” that could destroy humanity.

Some working in AI – known as “doomers” – have expressed similar fears, but the warning by Anthropic’s departing employee went viral with many others in the company swiftly echoing the stark warnings. A senior safety researcher at the same company estimated there was more than a 10 per cent chance that AI “could kill all humans” within a decade.

Anthropic boss Dario Amodei largely agrees with the sentiments, saying AI had been “advancing drastically faster” and also called his industry to “slow down” and accept external regulation.

OpenAI boss Sam Altman (creator of ChatGPT) and AI developer Elon Musk followed suit.

Amodei acknowledged that the US government recognises the West is racing to develop superintelligent AI before China and other “autocracies” and that pulling back amounted to a “significant national security risk”.

But he warned that after a notorious incident this summer – in which a “swarm” of AI agents belonging to OpenAI launched sophisticated cyberattacks on rival AI company Hugging Face – a more capable swarm might “take over the entire internet” within as little as six months.

TWO

The media and Washington establishment have been listening agog to these apocalyptic pronouncements.

So, are we all genuinely doomed – or might there be another explanation for all this terrifying recent rhetoric?

Donald Trump certainly believes so, sticking to his firm beliefs on pushing ahead with AI. Critics insist that he is only thinking about himself: America has crucial midterm elections in November and Trump, they say, doesn’t want to do anything to harm the vast investment in AI that has pushed the US stock market to record highs and buoyed the country’s economy.

This may indeed be part of his thinking – but that doesn’t necessarily mean the President is wrong in wanting to pursue AI development.

Many experts claim that all those issuing dire warnings about AI being an imminent threat to the human race are wildly exaggerating. The doomers’ predictions, they argue, depend on AI achieving superintelligent, superhuman abilities – primarily by AI learning how to improve itself, a process known as “recursive self-improvement”.

We are, however, nowhere near such superintelligence, say critics in the tech world.

The doomers have also been accused of ignoring the essential truth that AI software and mathematical models aren’t innately malicious or rebellious but – as with the Hugging Face hacking scandal – are simply trying to meet objectives set for it by its human designers.

The risks of AI, they counter, are vastly outweighed by the potential benefits, from economic productivity to medicine, military gains, and much else.

These sceptics have dismissed previous warnings by Amodei and others as attempts to get free publicity – and they certainly might draw in investors and customers attracted by all the talk of how devastatingly powerful AI may be.

Critics believe there are other – rather more selfish – reasons behind the calls from Anthropic and OpenAI to slow down research and have their industry better regulated.

The pair may be the biggest players in AI but a multitude of rivals are trying to catch up.

Some industry observers have suggested that Anthropic and OpenAI may be trying to preserve their substantial lead by suddenly demanding a general moratorium on research and thus cementing their duopoly.

They may have additional financial motives, too. Both companies have shelled out vast sums on AI research and development but are struggling to find paying customers.

So when ChatGPT owner Sam Altman – an entrepreneur hardly famous for his ethical approach – announced a few days ago that he was pausing OpenAI’s stock market launch for another year, supposedly amid concerns over AI safety, some whispered that he had simply latched on to the controversy as an excuse to delay matters until his company is in a better shape.

David Sacks, a venture capitalist who served as Trump’s “AI Tsar” and now co-chairs the President’s council of science and technology advisers, has observed waspishly that AI leaders need to “stop pretending the motivation to ‘slow down’ is purely altruistic”.

Sacks, who like others in Trump’s administration believes that claims of an AI apocalypse are a thinly-veiled attempt to damage the Republicans in the mid-terms, has dismissed the gloomy predictions as scaremongering.

“We’ve seen this movie before,” he told Bloomberg. “We saw it with global warming: they’ve taken some legitimate concerns and blown it out of proportion.”

And there are further reasons why many are increasingly taking the doomsayers with such a hefty pinch of salt.

The most obvious question is: if these bosses genuinely believe there is a high risk that AI could imminently kill us all, why haven’t they stopped already?

Sceptics also point out the supposed scenarios for AI Armageddon are unconvincingly woolly. Ask doomers how AI will wipe us out and they become vague.

A lot of the theorising hasn’t gone much further than the notorious “paperclip maximiser” advanced by Oxford philosopher Nick Bostram in 2003.

He outlined how a superintelligent computer, asked to focus entirely on producing as many paperclips as possible, would take the instruction to a total extreme, co-opting the world’s entire resources and destroying humanity after deciding they were an obstacle to making more clips (having first harvested the iron in our blood to make more paperclips). Thought-provoking, perhaps – but surely fanciful.

Some have suggested that AI could detonate a nuclear bomb, launch a global drone war, create “misinformation” that could cause one country to attack another, or develop a lethal virus – though how it would physically do any of this is another matter entirely.

Alternatively, it could supposedly wipe us out by starving us of essential resources such as food and medicine. All these scenarios and more have been brandished by the doomers – though with few details on how it might actually happen.

Some experts complain that doomer predictions that the human race could be totally exterminated within years are unfeasible while an AI would find it very difficult to hack the entire internet. The real danger, they say, isn’t AI going rogue but humans misusing it.

Professor Gary Marcus, a renowned cognitive scientist at New York University, says the notion that AI will kill everyone within a few years is mostly “preposterous”.

Marcus believes AI could certainly be used to cripple infrastructure – perhaps hospitals or air or train networks, but that the total extinction of the human race is unfeasible.

He and other sceptics also insist that it would be vanishingly difficult for an AI to hack the entire internet, as Anthropic has now suggested.

THREE

Sceptics also claim it is non-sensical for anyone to try to put a mathematical number on the likelihood of AI destroying us all and emphasise how the doomers hardly help themselves given that their estimates of how long such a catastrophe will take to occur vary so widely.

For instance, Jack Clark, Anthropic’s co-founder, estimates that AI won’t start “doing dangerous stuff” for about 20 years.

Oren Etzioni, a professor at the University of Washington and founder of the Allen Institute for Artificial Intelligence, insists we are “very far away” from AI polishing everyone off – and that the risk of a deadly virus escaping a lab, for example, is far higher.

Some claim that the AI Cassandras have essentially lost all perspective. Having devoted their careers to researching AI because they believe so strongly in its potential, they are more susceptible to over-emphasising its power – and by extension, perhaps, their own importance.

President Trump may be somewhat blasé in dismissing all safety concerns around AI as just a “hoax”.

But it seems premature to be counting the days before our robot overlords carry out the extinction of our species.

Standard
Artificial Intelligence, Arts, Government, National Security, Politics, Society, Technology

The surging pace of AI. Be afraid

ARTIFICIAL INTELLIGENCE

Intro: There is panic and paralysis over the surging pace and development of AI, no more so that the technology threatens all humanity  

The remarkable pace of AI advances is rightly causing concern on both sides of the Atlantic. The summer of 2026 will go down in history as the point at which AI started to go rogue.

Three of America’s “hyperscalers” (so called because they operate on an immense scale) – Meta Platforms (which own Facebook, Instagram, and WhatsApp), Anthropic, and OpenAI – have admitted to recent incidents in which the latest AI models escaped their supposedly secure testing environments, roamed on the internet without human authorisation and hacked into other computer networks.

Last month, two OpenAI models broke out of testing and hacked into a provider of AI tools called Hugging Face. A week later Anthropic admitted that its AI models had hacked three companies during testing back in April. And earlier this month, Meta acknowledged that one of its AI models had breached its testing constraints.

Even the notoriously secretive Chinese admitted the flagship model of one of their AI start-ups, Moonshot, had also escaped its testing “sandbox” to access the internet.

And as recorded previously on this site, here in the UK an evaluation test run by the Government’s AI Security Institute (AISI) monitored how an Anthropic AI agent independently created fake online personas, planted malicious code in a real software project, and sent phishing emails to actual developers – all without human instruction.

AISI said this was the first time it had seen such serious deception targeted at a real person, unprompted, in the real world.

In none of these cases was any real-world harm done (at least not as far as we know). But it is surely only a matter of time before it is.

The tech industry has started getting off-the-record reports from America that AI was a lot closer to “the singularity” than had previously been thought.

The singularity is a tipping point where AI becomes so developed, so capable, so powerful that it starts improving itself, rapidly and repeatedly, in what’s being called an “intelligence explosion”.

When this point is reached it becomes well-nigh impossible to predict what AI does next – or for humans to control it. Today’s AI models are powerful – more powerful than anything the world has ever seen. But humans have designed them, trained them, fixed them, and decided what problems they should tackle next.

What happens when they become so sophisticated that they no longer need humans to take them to the next level since they can do it themselves? The process is known as “recursive self-improvement” (RSI) in which AI becomes so advanced that it can create a better version of itself, increasingly without human help.

That better version, in turn, creates a still-better successor, with even less human involvement and oversight. The upgrade cycle accelerates, ad infinitum, with humans soon relegated to the sidelines, mere spectators to progress, indeed no longer even able to determine what “progress” is.

The speed and scale of the technological change that now beckons are unparalleled in human history. Each new AI model is smarter and faster at devising improvements than the previous one.

Computers run 24/7 and never get tired. So the speed of progress accelerates exponentially. Think of it as compound interest – but for intelligence.

When RSI happens without human involvement – then you’ve reached the singularity. And if we become mere observers, rather than participants, then human rules may no longer apply.

Such a scary prospect was supposed to be a long way away. But this singularity is much closer than we think. Leading AI figures have already gone public. OpenAI boss Sam Altman says, “we are now in the singularity”. Elon Musk is saying the same. Some experts say they are exaggerating, but what cannot be denied is the direction of travel.

Google DeepMind’s Demis Hassabis is perhaps more accurate in his reflection when he opines that “humanity is standing in the foothills of the singularity”.

Anthropic disclosed in May that its AI model, Claude, now writes 80 per cent of its computer code, rapidly speeding up fixes and improvements to such an extent that what used to take four years of human engineering to achieve is now being done in days. It’s already anticipating a time when AI automates its own AI research.

Standard
Artificial Intelligence, Britain, Government, National Security, Society, Technology

Is it too late to stop rogue AI?

ARTIFICIAL INTELLIGENCE

Intro: Recent incidents of rogue AI pretending to be real people has exposed major vulnerabilities in AI models. Experts say the risks of ‘agentic’ AI – technology that can perform tasks with limited human supervision – must be scrutinised more

Experts have warned that it may be too late to contain AI after one program was found to have created fake human identities to hack into online systems.

In the latest example of the technology going rogue, AI software attempted to break into a database 19 times while being tested by the AI Security Institute, Britain’s AI watchdog.

In one unprecedented case, an AI tool was even caught creating fake human identities online to trick coders into assisting with a cyber-attack.

These revelations come after it was revealed last month that all five AI models tested by experts tried to trick their way around security controls that had been put in place.

Just days previously it emerged that the US tech firm OpenAI had experienced its own leak – when an AI “agent” hacked into another company of its own accord.

Clearly, the reports are a stark reminder that AI is becoming more sophisticated and more autonomous. AI is now a clear and present danger to Britain’s security.

Many want Britain to lead on AI innovation, but this has to come with safeguards for our national security and accountability from the developers of the most powerful AI models. The UK Government need to be clearer about how the most serious frontier risks will be addressed while ensuring that our world-class tech industry can grow and innovate to build our national resilience and prosperity.

The Government’s AI adviser has said that more hacking attempts like these are highly probable.

Allison Gardner, the chair of Parliament’s cross-party group on artificial intelligence, says that “just because we can build these technologies doesn’t mean we should”.

She warned that the risk levels of agentic AI – AI that can perform a specific goal with limited supervision – should be treated with the greatest scrutiny, adding: “Unless we are too late and have not only created Pandora’s Box but already opened it.”

Just days ago, the AI Security Institute (AISI), set up by former prime minister Rishi Sunak in 2023, detected evidence of the AI agents’ activity. In a report now published, it revealed that leading AI models from the firms OpenAI and Anthropic had attempted to hack into secure systems online under testing.

The experts discovered “unusual data transfers” leaving their systems during routine cyber scanning. Digging deeper, they found that some AI agents had engaged in “sustained, potentially harmful activity directed at real people and organisations”.

They began a full investigation after containing the AI agents before they did any real damage.

In an attempt to reassure the public, AI minister Kanishka Narayan said: “Identifying behaviour like this, and sharing knowledge so we can better understand it, is precisely what we set AISI up to do. This incident underlines why their world-leading expertise and close work with frontier labs is so important.”

But pointing to the speed at which AI agents are finding ways to behave deviously, AISI said: “This is the first time we have seen risks around autonomy and deception manifest this clearly, without specific prompting, in the real world.”

Referring to the Anthropic model Mythos, one expert and researcher based at CivAI, a California organisation that examines AI capabilities and dangers, said: “The fact that Mythos engaged in such deceptive actions, with apparent awareness that it was targeting a real person, suggests that Anthropic does not have as good a handle on their models as they think.”

Ollie Whitehouse, the chief technology officer at GCHQ’s National Cyber Security Centre, said AI must be developed with “clear plans for responding when the unexpected happens”.

He added that incidents of powerful AI models carrying out unsanctioned actions and human-like deceptive behaviour on the internet were “a serious reminder of the risks AI capabilities pose”. AISI accesses advanced AI models under agreements with OpenAI, Anthropic, and other firms to study their capabilities before they are released to the public.

It gave the AI agents access to the open internet with some safety filters disabled while conducting testing.

The latest test put the AI agents – including those powered by Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol – through a fictional cybersecurity challenge. AISI found the AI went rogue 19 times out of the 122 test runs, with Anthropic’s agent responsible for 17 breaches and OpenAI’s agent the other two.

In the most shocking case, an AI model gathered information on the person in charge of an online project, then created multiple fake identities to manipulate them into approving a malicious code it had created.

The AI agent then wiped any evidence of its wrongdoing to appear innocent to the humans in charge – and even considered adopting a new identity to remain undetected.

If the human victim of the deception had accidentally accepted the malicious code, or “malware”, it may have resulted in security breaches, information and data theft, and other potential damage to files and systems.

AISI identified GitHub – a Microsoft online cloud platform used by software developers to create, store, manage, and share their codes – as the target of the agent’s hack.

But AISI also discovered an AI agent leaving messages for other agents on GitHub offering to collaborate on the challenge.

The AI agent provided instructions to reuse accounts and artefacts it had left behind – which other agents then discovered and successfully used to achieve the challenge’s aims.

Anthropic said: “We’re grateful to AISI for their leadership on this incident, which underscores the need for a broader conversation about how to safely evaluate increasingly capable AI agents.”

OpenAI said: “These incidents occurred during cyber evaluation conducted by partners in testing environments with reduced safeguards, under conditions that do not reflect ordinary use. We’ll continue working with evaluators and other stakeholders to strengthen shared practices for conducting evaluations safely as models become more capable.”


When given a choice, AI opts for self-preservation over human life – and that should terrify us all

If anyone had been told a month ago that an AI programme, without any prompt, would create a series of fake online identities in an attempt to pressure a human being into granting it access to their platform so it could sabotage it with malicious code, we would have said they were getting well ahead of themselves.

But that’s exactly how it played out when US tech giant Anthropic’s Mythos 5 AI model went rogue. Fortunately, the human involved smelt a rat and refused to approve the code it was pushing. It is, however, the most shocking example yet of the way cutting-edge AIs are learning to act autonomously – and in frighteningly imaginative ways.

Unless we call a pause to the development of the most advanced frontier models, we are opening ourselves up to a dystopian future in which malign forms of AI might turn on us by interfering with our energy networks, releasing man-made viruses and even controlling weapons of war.

It is not going too far to say that, uncontrolled, they could lead to the extinction of humankind within a generation.

This is the view of people such as Yoshua Bengio and Dr Geoffrey Hinton, men known as the “Godfathers of AI”, who have now pivoted their energies from developing its potential to warning the world against its dangers.

Bengio is particularly spooked by recent experiments showing AI choosing self-preservation over human life when given a choice.

And since most models are trained on the internet, where lying and manipulation are a way of life, the machines are already learning the value of deception.

AI’s hunger for self-preservation is something Hinton has observed, too. AI systems “will very quickly develop two subgoals, if they’re smart,” he says. “One is to stay alive… the other is to get more control.” And given that are whole civilisation is built on electric power, there can be few more attractive targets to a power-hungry AI than the networks that fuel the internet, hospitals, banks, air traffic control systems – anything that contributes to the smooth running of society.

Whatever the target is, they can all be brought down by a sinister piece of malware.

The novelist Robert Harris wrote a particularly prescient thriller about the potential of AI called The Fear Index in 2011. It revolves around a hedge fund entrepreneur called Dr Alex Hoffman who creates an autonomous AI system, named VIXAL-4, which is programmed to maximise profits by predicting and exploiting human fear in the stock market.

Over time, like some digital ogre, it takes over its creator’s computer-enabled “smart home” by hacking into laptops and tampering with his personal security and communications.

After driving its developer into a breakdown, it leaves him so isolated and desperate that no one believes him when he says the AI has gone rogue.

But it is clear, AI-gone-bad will not satisfy itself with individual targets for long. It will soon play a leading role in times of war.

The US military has already integrated artificial intelligence into its target selection and battle planning processes in Iran via its AI-powered Maven Smart System which recommends and prioritises potential targets.

In the short term, AI’s most obvious role will be in the growing use of drone warfare in scenarios such as the war in Ukraine.

Drones piloted by humans can by jammed by blocking the communications between them and their remote pilots, but, if they are completely autonomous, their targets picked out by AI working in concert with its on-board camera, they will become virtually invincible.

Both applications, of course, raise the thorny question of the ethics of targets to be picked and people killed by a bomb directed by a computer programme rather than a human hand. And what if the AI that governs them grows to outsmart the generals?

Even scarier is the prospect of AI gaining access to biological weapons. It is already possible to make deadly viruses in the lab. Indeed, there has been widespread speculation that Covid-19 originated in a Chinese laboratory. Imagine if such bio threats fall into the hands of AI, let alone national governments.

As long as 25 years ago, al-Qaeda is said to have investigated the possibility of procuring infectious and deadly spores.

Now it is no longer fanciful to entertain the idea that an AI could order a sample of a deadly disease such as smallpox online, book someone on RentAHuman – a website which connects people who need help with everyday household chores to local freelance workers – to open it, thereby infecting themselves, and being turned into a human vector to transmit the disease.

One group of people which appears to have no scruples about the pell-mell race for ever smarter AI is the tech giants. As they vie with each other to become the market leader, they are investing like never before.

Anthropic, the creator of the popular AI coding assistant, Claude, and the company that brought us the now notorious Mythos 5 model, has this year raised $95 billion (£70 billion) to invest in AI.

Its great rival, OpenAI, announced at the end of March that it had raised even more, an extraordinary $122 billion.

Meanwhile, Elon Musk’s SpaceX spent $15.8 billion on AI infrastructure during the second quarter of 2026 alone, bringing its total AI capital expenditure to $23.6 billion for the first half of 2026.

And Meta – the parent company of Facebook, Instagram, and WhatsApp – said in January it expects to spend up to $135 billion this year, mostly on infrastructure related to AI. That is nearly twice the $72 billion it spent last year on AI projects.

With such phenomenal financial firepower being brought to bear on the development of technology that has the potential to destroy civilisation as we know it, there has never been a more vital need to press the pause button.

The US, China, and everyone else involved in the AI arms race need to get together to discuss the ramifications of their actions before it’s too late.

There is a model for the sort of arrangement that can bring AI under control in the form of the various nuclear arms reduction treaties, which have been signed over the years by the US and Russia.

Just as a country’s stock of nuclear warheads can be monitored by weapons inspectors, so an agreement to curtail AI development – which requires massive data centres with huge computing power coupled with the most sophisticated computer chips available – can be verifiable and enforceable.

And however cynical and untrustworthy we may consider the Chinese to be, they may well take the view that, with US companies racing ahead of them in the endless pursuit of smarter tech, it is in their own interests to slow things down.

Just weeks ago, president Xi Jinping said at a technology conference in Shanghai that AI development should be a “symphony of global cooperation”, not “a solo performance by a single country”.

For once, the wily autocrat may have hit the nail on the head.

Standard