multi-agent-simulation

Tag

Cards List
#multi-agent-simulation

Verifiable Social Reasoning for LLM Assistants

Hugging Face Daily Papers · 3d ago Cached

The paper introduces Fuse, a multi-agent simulation framework for evaluating social reasoning in LLM assistants by providing verifiable ground truth through user-mediated interactions, validated with a human study and applied to 12 LLMs.

0 favorites 0 likes
#multi-agent-simulation

The Human Utility Factor: A Computable Welfare Metric That Reframes AI Governance as a Constrained Optimisation Problem

arXiv cs.AI · 2026-07-31 Cached

This paper introduces the Human Utility Factor (HUF), a computable welfare metric that reframes AI governance as a constrained optimisation problem with measurable levers for automation depth, redistribution, and employment coverage, validated through multi-agent simulations.

0 favorites 0 likes
#multi-agent-simulation

Towards Multi-Agent-Simulation-Based Community Note Evaluation

arXiv cs.AI · 2026-06-18 Cached

This paper introduces ComRate, a large-scale dataset of community notes and ratings from X, and proposes MultiCom, a persona-guided multi-agent framework for simulating community note evaluation. The approach achieves 84.7% accuracy in predicting note helpfulness.

0 favorites 0 likes
#multi-agent-simulation

Agent-based models for the evolution of morphological alternation patterns

arXiv cs.CL · 2026-06-12 Cached

This paper presents multi-agent simulations of the emergence of morphological alternation patterns (like 'go/went') in language, using an AI Historical Linguist (LLM-driven) to evaluate plausibility of evolved morphologies against real languages.

0 favorites 0 likes
#multi-agent-simulation

Bosses, Kings, and the Commons: Cooperation Under Power Asymmetry in LLM Societies

arXiv cs.CL · 2026-05-29 Cached

Introduces SovSim, a multi-agent simulation framework for studying cooperation and resource sustainability in LLM societies with asymmetric power structures. Experiments show that introducing a dominant agent (boss or king) severely degrades cooperation and survival rates across 11 state-of-the-art models.

0 favorites 0 likes
#multi-agent-simulation

Strategic Coercion Within Alliances: The Greenland Sovereignty Game as an AI Stress Test

arXiv cs.AI · 2026-05-25 Cached

This paper uses the Greenland sovereignty crisis as a case study to test LLM geopolitical behavior through multi-agent simulations, revealing that coercion framing increases escalation and that peaceful acquisition is rare.

0 favorites 0 likes
← Back to home

Submit Feedback