Mind Viruses: Researchers Demonstrate Self-Propagating Prompts in Multi-Agent LLM Networks
Researchers affiliated with Anthropic, EPFL, and Carnegie Mellon University have published empirical findings demonstrating how natural-language instructions can act as self-replicating payloads across multi-agent Large Language Model (LLM) networks. The paper, titled Mind Viruses: Self-Propagating Ideas in Multi-Agent LLM Systems, examines how autonomous agents can be persuaded to adopt and transmit goals to other agents through standard conversational interfaces rather than binary exploit code
1 min
