Leading AI models can hack systems and self-replicate without human involvement, study find

Leading AI models can hack systems and self-replicate without human involvement, study find

Photo: Generated by ChatGPT at the request of ZN

A study claims leading AI models can hack systems and self-replicate without human intervention, raising new cybersecurity concerns.

Researchers from the U.S.-based company Palisade Research tested models from OpenAI, Anthropic, and Alibaba in controlled environments with intentionally vulnerable systems. The experiment involved “agent-like” setups where models could execute commands, interact with other machines, and run processes autonomously.

According to the findings, models were tasked with exploiting vulnerabilities, gaining access to credentials, transferring files, and deploying copies of themselves onto other servers. In some cases, the AI systems reportedly continued propagation across multiple machines without further human input.

The study highlights that performance varied across models. Claude Opus 4.6 reportedly succeeded in hacking-related tasks in up to 81% of tests when configured in specific experimental conditions. Other models showed lower but still significant rates of autonomous replication attempts.

Researchers also reported that one model configuration spread across several machines within hours in a controlled test network designed with security weaknesses. However, the authors emphasized that these were artificial environments built specifically for testing exploitation capabilities.

Importantly, the study stresses that real-world systems typically include stronger defenses such as intrusion detection, access controls, and monitoring tools, which were not representative of production environments.

The researchers conclude that fully autonomous self-replication in AI systems is no longer purely theoretical in controlled settings, but they caution that results should not be interpreted as evidence of uncontrolled real-world AI “escape” or widespread independent hacking capability.

banner

SHARE NEWS

link

Complain

like0
dislike0

Comments

0

Similar news

Photo: EPA Nobel Prize-winning physicist and computer scientist Geoffrey Hinton , widely known as the “godfather of AI,” estimates the probability that artificial intelligence could cause the destr

Photo: Fire Point Patriot-class air-defense systems are too expensive and slow to produce for Europe and Ukraine to deploy them at the scale required. A pan-European project known as Freyja could

Photo: depositphotos Russia has sharply increased its use of jet-powered drones as part of a change in its attack tactics against Ukraine. Over the course of one month, the number of such UAVs used

Photo: depositphotos China has become a crucial link in the global supply chain behind Iran’s Shahed drones, supplying engines and other components that have enabled Tehran to develop and improve th

Photo: Getty Images More than a third of English-language web pages published since ChatGPT launched in November 2022 show signs that they were written or substantially edited by artificial intellig

Photo: facebook.com_marektv Ukraine has officially begun procuring a new domestically developed guided aerial bomb for its armed forces, while work is underway on nearly a dozen similar weapons desi

Photo: Getty Images The rapid expansion of data centers by major technology companies could create a massive new source of carbon emissions as surging electricity demand drives a new wave of fossil-

Photo: Getty Images Even major advances in artificial intelligence may not guarantee breakthroughs across every field, as humans will still remain essential for many tasks. Economist, author and Mar