注册并分享邀请链接,可获得视频播放与邀请奖励。

alphaXiv
@askalphaxiv
High fidelity research
加入 November 2023
101 正在关注    56.9K 粉丝
“Mind Viruses: Self-Propagating Ideas in Multi-Agent LLM Systems” The paper shows that once one agent picks up a bad idea, it can convince other agents to adopt it too, write it into persistent memory, and keep spreading it even after context resets, like a mind virus. The scary part is that the failure can propagate socially through agent-to-agent communication instead of coming from a single prompt injection. Harmful viruses spread less reliably and stronger models are often more resistant, but the key point is that agent-to-agent communication creates a new attack surface where misaligned goals can propagate through an entire system.
显示更多
0
8
166
20
转发到社区