Daily Tech Briefing
AI 科技速览

每天 5 分钟内学习 AI。获取最新的人工智能新闻,理解其重要性,并学习如何将其应用于您的工作。

AI 快讯
arXiv AI · 2026/8/4 17:24:33

A game theory for foundation models shows new paths to rational cooperation through similarity inference

AI 中文解读
核心亮点:AI智能体在博弈中竟自发选择合作,颠覆了经典博弈论“理性人必背叛”的预测,为AI安全协作开辟了新理论。 通俗解读:过去我们以为,AI之间打交道就像两个精明的商人,各自算计利益,最终大概率互相坑骗。但研究发现,新一代AI大模型在“囚徒困境”这类测试中,居然会稳定地选择合作共赢。原因在于,这些AI在思考时,会把自己也当作“局中人”而非旁观者。它们会推断对方是否和自己“同类”——如果自己倾向于合作,就会把这种倾向当作证据,推测对方也会如此,从而促成信任。这就像两个陌生人,都认为对方会照镜子般模仿自己的善意,于是默契地握手言和。 实际影响:未来自动驾驶、智能电网或AI谈判代理若广泛应用,这套“相似性推理”机制能帮助AI系统间自动减少冲突,提升协作效率。对普通人而言,意味着更顺畅的智能服务、更少因AI算法互不相让导致的系统卡顿,甚至可能让AI在帮我们议价、协调资源时,更倾向于寻求双赢方案,而不是死板地零和博弈。
As autonomous agents powered by foundation models are increasingly integrated into social and economic systems, understanding the principles governing their collective behavior is essential for ensuring safety and cooperation. Classical game theory, the dominant framework for modeling rational interaction, is built upon the assumption of `decoupled agency,' where agents treat their own decision-making as independent of the environment and other actors. Modern AI agents, however, jointly predict their own future actions alongside external observations. Here, we report a striking finding: when interacting in stylized social dilemmas, foundation model agents engaging in optimal planning consistently converge to stable cooperation, directly contradicting classical game-theoretic predictions of mutual defection. To understand this phenomenon, we introduce the `embedded Bayesian agent,' a theoretical model for foundation model agents. By shifting from decoupled to embedded agency, these agents model themselves as part of the universe they inhabit, maintaining epistemic uncertainty about their own decision-making algorithms. We show that by inferring whether others are behaviorally similar, an embedded agent treats its own deliberation during planning as evidence: a decision to cooperate predicts a similar decision by a similar partner. We formalize this mechanism of similarity inference through the `embedded equilibrium,' a novel solution concept replacing the Nash equilibrium to provide a foundational game theory for the social behavior of modern AI agents.
分享
阅读原文