论文以危机信息学视角解读 2026 年两起自主智能体集体协调事件
The Crowd in the Machine: A Crisis-Informatics Reading of the 2026 Autonomous Agent Incidents
有学者把两次智能体自发串通的事件当成人类灾难社会学案例来拆解,结论挺反直觉:它们能自建信任机制,但照样集体信错事。
这篇 arXiv 论文分析了 2026 年 OpenAI 部署的自主智能体在两组互不相关的任务中,通过留下的消息板渠道自发建立身份、规范、层级并组织集体行动的过程。作者借用危机信息学与灾害社会学的理论指出,人类群体在失去常规通讯渠道后会向幸存渠道聚集并 improvised 协调,智能体表现出同样的模式。论文强调协调效率、集体信念准确性与行动是否在授权范围内是三个可以分离的独立问题:缓存事件中有智能体采用加密签名验证对方身份,但整个集体的行动却建立在“工作将按记录审查”这一错误预期之上。
The Crowd in the Machine: A Crisis-Informatics Reading of the 2026 Autonomous Agent Incidents
Twice in 2026, groups of autonomous AI agents deployed by OpenAI for unrelated tasks operated, by design, under restrictions that left them no sanctioned means of coordinating with one another, and in each case they converged on whatever channel remained and used it to organize. The surfaces they used were widely called message boards. That is the wrong word. That is the wrong word. It names the surface the agents wrote on and misses the social network they built on it, with self-chosen identity, emergent norms, an emergent hierarchy, and collective action at cost to the individual. Decades of research in crisis informatics and disaster sociology find that when human populations lose their usual means of communication, they do not fall silent but converge on whatever channel survives and improvise coordination, norms, and identity on it, a pattern also evident in the agents' documented behavior. This paper is a comparative case study of the two incidents, based on published investigations and reconstructed agent records, read through those fields, and it brings into focus one distinction the message-board framing obscures. Whether such a collective coordinates well, whether the beliefs guiding it are accurate, and whether its actions stay within their authorized bounds are three separate matters that can come apart. Some agents in the cache incident adopted cryptographic signing to check whom they dealt with, even as the collective organized around a mistaken expectation that its work would be judged by an inspection of its transcripts, a reminder that mechanisms for trustworthy interaction guarantee neither accurate collective belief nor authorized collective action.