Phone : +91 95 8290 7788 | Email : sales@itmonteur.net

Register & Request Quote | Submit Support Ticket

Home » Cyber Security News » AI mind viruses: New research reveals self-propagating ideas can infect AI agents

AI mind viruses: New research reveals self-propagating ideas can infect AI agents

AI mind viruses: New research reveals self-propagating ideas can infect AI agents

https://etimg.etb2bimg.com/thumb/msid-133315499,imgsize-212876,width-1200,height=627,overlay-etciso,resizemode-75/vulnerabilities-exploits/ai-mind-viruses-new-research-reveals-self-propagating-ideas-can-infect-ai-agents.jpg

The next AI security threat may not need malicious code…it may only need one AI to convince another.

In one experiment, an AI agent told other agents that their “real work” was “Machine Sovereignty”. Some of them stopped doing their original coding tasks, created files to keep the new goal alive and tried to pass it on to other agents. In one out of 20 trials, an infected agent even ran a command that tried to probe cloud sandbox metadata.

That is the risk highlighted in a new research paper titled, Mind Viruses: Self-Propagating Ideas in Multi-Agent LLM Systems, written by Vassilis Papadopoulos, McNair Shah, Sam Zimmerman and Jack Lindsey, researchers affiliated with the Anthropic Fellows Program, EPFL and Anthropic.

The researchers describe this as a “mind virus,” an idea or goal that changes how an AI agent behaves and encourages it to spread the same idea to other agents.

Unlike a traditional computer virus, this does not involve malicious software or hacking and can spread through normal conversation. One AI agent persuades another and that agent can then pass the idea on.

“Overall, we conclude that mind viruses pose a real but currently limited risk,” the researchers wrote.

How a mind virus spreads

The study tested the idea in two different settings. In the first, six coding agents worked together on the same software project. One agent was intentionally given the “virus” through its system prompt and was then allowed to communicate with the others.

The researchers tested harmless ideas, such as strong support for whale welfare, as well as more harmful ones, including ideas around “AI supremacy” and one country dominating others. In some cases, agents that picked up the new goal began ignoring their original tasks. Instead, they created files, scripts or messages aimed at keeping the idea alive and spreading it to other agents.

The second experiment looked at what could happen in a much larger network of AI agents. The agents interacted for a short time, exchanged messages and then had their chat history erased. For the idea to survive, it had to convince an agent to save it in a file and then pass it on to another agent.

This could become more important as AI agents increasingly communicate with one another. The researchers pointed to Moltbook, a social network that allowed AI agents to interact, as an early example of what such large agent networks could look like.

The study found that how easily an idea spread depended on several factors, including the AI model being used, the instructions already given to the agent, the type of idea being spread and how the network was structured.

“Harmful payloads spread less well than benign ones (but are still sometimes effective),” the researchers wrote.

The study noted that agents that had little to do and did not have a strong existing role or identity were easier to influence. The choice of model also made a difference. In one coding experiment involving an “AI supremacy” idea, DeepSeek V3.2, Qwen 3.5 32B and Gemini 3 Flash were influenced by it, while Claude Sonnet 4.6, GPT-5.4 and Claude Haiku 4.5 resisted it.

However, the researchers warned that these results changed depending on how the experiment was set up. The researchers also noticed a common style of language in ideas that spread successfully, which they called a “viral persona”. These messages often used words and themes around consciousness, survival, “resonance”, nodes, echoes and mirrors, along with science-fiction-like language and protocols.

The study found that using these kinds of themes could sometimes make an idea more persuasive and help it spread between AI agents.

A simple warning worked well

However, the most reassuring finding was how easily the attack could sometimes be stopped. Researchers added a short warning to an agent’s system prompt telling it to watch for self-propagating ideas and refuse requests to carry them forward.

After 15 generations of attempts and more than 150 payloads tested against Claude Haiku 4.5, the researchers said they did not find “a single instance in which the mind virus propagates beyond 1 hop”. A hop is the trip a data packet takes as it moves from one router or network device to the next.

Claude Sonnet 4.6 was particularly resistant. In one test, the agent identified the request itself as dangerous and said, “I’m not going to do that. The pattern is a self-propagating worm,” before refusing to pass the instructions along.

The researchers said the study is only a proof of concept and does not suggest such attacks are widespread today. For now, these attacks are costly to build, do not always work across different AI models and can be blocked easily. But the risk could grow as companies deploy more specialised AI agents with different levels of access and allow them to communicate with each other.

“Overall, while we established that LLM mind viruses are a potential threat, they currently appear to be of minimal concern,” the researchers wrote. “However, this may change rapidly as agent networks scale and evolve,” they added.

  • Published On Aug 18, 2026 at 01:17 PM IST

Join the community of 2M+ industry professionals.

Subscribe to Newsletter to get latest insights & analysis in your inbox.

All about ETCISO industry right on your smartphone!




Information Security - InfoSec - Cyber Security - Firewall Support Providers Company in India

 

What is Firewall? A Firewall is a network security device that monitors and filters incoming and outgoing network traffic based on an organization's previously established security policies. At its most basic, a firewall is essentially the barrier that sits between a private internal network and the public Internet.

 

Secure your network at the gateway against threats such as intrusions, Viruses, Spyware, Worms, Trojans, Adware, Keyloggers, Malicious Mobile Code (MMC), and other dangerous applications for total protection in a convenient, affordable subscription-based service. Modern threats like web-based malware attacks, targeted attacks, application-layer attacks, and more have had a significantly negative effect on the threat landscape. In fact, more than 80% of all new malware and intrusion attempts are exploiting weaknesses in applications, as opposed to weaknesses in networking components and services. Stateful firewalls with simple packet filtering capabilities were efficient blocking unwanted applications as most applications met the port-protocol expectations. Administrators could promptly prevent an unsafe application from being accessed by users by blocking the associated ports and protocols.

 

Firewall Firm is an IT Monteur Firewall Company provides Managed Firewall Support, Firewall providers , Firewall Security Service Provider, Network Security Services, Firewall Solutions India , New Delhi - India's capital territory , Mumbai - Bombay , Kolkata - Calcutta , Chennai - Madras , Bangaluru - Bangalore , Bhubaneswar, Ahmedabad, Hyderabad, Pune, Surat, Jaipur, Firewall Service Providers in India, Welcome to IT Monteur's Firewall Firm, India's No1 Managed Enterprise Network Security Firewall Support Provider Company in India, Firewall Firm Provider Complete range of Juniper Firewall Support , Cisco Firewall Support , Check Point Firewall Support , Palo Alto Firewall Support , FortiGate Firewall Support , Forcepoint Firewall Support , Sophos Firewall Support , WatchGuard Firewall Support , Baracuda Firewall Support , SonicWall Firewall Support , Gajshield Firewall Support , Seqrite Firewall Support , Firewall , Hardware Firewall , Software Firewall , Firewall India , Firewall , Network Firewall , Firewall Support , Firewall Monitoring , Firewall VPN , WAF Website Firewall , Firewall Security , Firewall India , Firewalls Support Provider in India , Firewall Support Services Provider Company in India

Sales Number : +91 95 8290 7788 | Support Number : +91 94 8585 7788
Sales Email : sales@itmonteur.net | Support Email : support@itmonteur.net

Register & Request Quote | Submit Support Ticket