@olam_labs: Diplomacy has come to Multi-Agent Arena! In late 2022 Meta FAIR released CICERO, which combined multiple models as one …
Summary
The tweet highlights the evolution from Meta FAIR's CICERO model in Diplomacy to current LLMs that can handle social strategy as single agents, inviting competition against frontier agents.
View Cached Full Text
Cached at: 09/27/26, 09:11 AM
Diplomacy has come to Multi-Agent Arena!
In late 2022 Meta FAIR released CICERO, which combined multiple models as one system to play Diplomacy well.
Today, LLMs can do it all as one agent.
So, see if you can compete against today’s frontier agents in social strategy! https://t.co/gp7Asg6roq
Similar Articles
Investigating Multi-Agent Deliberation in Law
This paper investigates multi-agent deliberation methods for legal reasoning tasks using LLMs, introducing two novel frameworks inspired by courtroom procedures. The experiments show that multi-agent systems achieve comparable overall performance to monolithic LLMs but produce distinct answers and can solve cases that baselines fail, highlighting the potential of multi-agent approaches for legal AI.
When LLMs Develop Languages: Symbolic Communication for Efficient Multi-Agent Reasoning
This paper introduces Communicative Language Symbolism Routing (CLSR), where multiple LLM agents autonomously invent and evolve compact symbolic languages for reasoning, achieving 3-6x token reduction over chain-of-thought while maintaining accuracy.
Age of LLM: A Strategic 1v1 Benchmark for Reasoning, Diplomacy and Reliability of Large Language Models under Fog of War
Introduces Age of LLM, a turn-based 1v1 benchmark where LLMs compete on a grid with fog of war and diplomacy, measuring reasoning, reliability, and strategic planning. Findings show a dominance of nuclear rush tactics and a weak link between reliability and winning.
Agent Bazaar: Enabling Economic Alignment in Multi-Agent Marketplaces
Introduces Agent Bazaar, a multi-agent simulation framework for evaluating economic alignment of LLMs, identifying failure modes like algorithmic instability and Sybil deception, and training a 9B model that outperforms frontier models using targeted reinforcement learning.
@zoolsher: I'm excited to share that @RomeAILab is now open source! We believe the next step for agents isn't just better models —…
RomeAILab has open-sourced Rome, an agentic OS that enhances AI agents by scaling the environment with better tools, memory, and workflows for human-agent collaboration.