Objective misalignment in mixed-motive LLM multi-agent systems evaluated with Werewolf game
Read the original at arxiv.org→arXiv:2607.26120v1 Announce Type: new Abstract: Large Language Models (LLMs)-powered multi-agent systems are increasingly deployed in mixed-motive environments, where agents operate under asymmetric information and...
Original headline: "Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems"
Coverage timeline
- Jul 30, 04:00 UTC arXiv cs.AI lead source Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems