multi-agent-reinforcement-learning

Tag

Cards List
#multi-agent-reinforcement-learning

Evolutionary Stability Does Not Guarantee Learning Accessibility: A Multi-Agent Reinforcement Learning Perspective on Cooperation Emergence

arXiv cs.AI ↗ · 6d ago Cached

This paper examines why evolutionarily stable cooperative outcomes in multi-agent systems may not be reachable through decentralized reinforcement learning algorithms, demonstrating distinct properties between stability and learning accessibility.

0 favorites 0 likes
#multi-agent-reinforcement-learning

Fully Byzantine-Resilient Multi-Agent Reinforcement Learning

arXiv cs.LG ↗ · 2026-09-23 Cached

The paper proposes FRAC-MARL, a decentralized actor-critic multi-agent reinforcement learning method that achieves full Byzantine resilience by leveraging redundancy in communication, ensuring convergence to optimal parameters even under adversarial attacks.

0 favorites 0 likes
#multi-agent-reinforcement-learning

DualSQL: Text-to-SQL with Multi-Agent Reinforcement Learning

arXiv cs.CL ↗ · 2026-09-17 Cached

DualSQL proposes a multi-agent reinforcement learning framework for Text-to-SQL, using a single shared model to jointly optimize schema linking and SQL generation, achieving state-of-the-art accuracy with smaller model sizes.

0 favorites 0 likes
#multi-agent-reinforcement-learning

LLM-Enhanced Multi-Agent Reinforcement Learning for Unified Electric Vehicles-Charging Station-Grid Optimization in Public Charging Systems

arXiv cs.AI ↗ · 2026-09-15 Cached

This paper proposes an LLM-enhanced multi-agent reinforcement learning framework to simultaneously optimize electric vehicle charging scheduling, station profitability, and grid stability, using LLMs for feature selection and adaptive weighting, outperforming state-of-the-art methods with reduced training time.

0 favorites 0 likes
#multi-agent-reinforcement-learning

JaxAHT: A JAX-Based Library for Ad Hoc Teamwork

arXiv cs.AI ↗ · 2026-09-15 Cached

JaxAHT is an open-source JAX-based library that accelerates and standardizes Ad Hoc Teamwork research, providing a unified framework for teammate generation, training, and evaluation with significant performance improvements and a suite of evaluation teammates.

0 favorites 0 likes
#multi-agent-reinforcement-learning

DRG-MAPPO: Hierarchical Dynamic Role-Graph Multi-Agent Reinforcement Learning for Cooperative Air Combat

Hugging Face Daily Papers ↗ · 2026-09-10 Cached

A hierarchical multi-agent reinforcement learning framework combining graph attention and dynamic role assignment improves tactical coordination and win rates in air combat.

0 favorites 0 likes
#multi-agent-reinforcement-learning

The Role of Network Topology and Opponent Information in Shaping Cooperation in Multi-Agent Reinforcement Learning Systems

arXiv cs.AI ↗ · 2026-09-01 Cached

This paper explores the impact of network topology and opponent information on the emergence of cooperation in multi-agent reinforcement learning systems, specifically in the Iterated Prisoner's Dilemma, finding that graph structure and information availability significantly influence cooperative strategies.

0 favorites 0 likes
#multi-agent-reinforcement-learning

SIGMA: Structured Noise-Effect-Aware Grouped Multi-Agent Aggregation

arXiv cs.AI ↗ · 2026-08-28 Cached

This paper proposes SIGMA, a hierarchical collaboration framework for cooperative multi-agent reinforcement learning that learns robust representations under noisy observations by exploiting cooperation structures through density-based grouping and aggregation methods.

0 favorites 0 likes
#multi-agent-reinforcement-learning

Decentralized Multi-Player Q-Learning in Episodic Markov Decision Processes with Information Asymmetry

arXiv cs.LG ↗ · 2026-08-14 Cached

This paper studies decentralized multi-player Q-learning in episodic Markov decision processes under three forms of information asymmetry, proposing algorithms that achieve regret bounds matching the single-agent Q-learning rate up to logarithmic factors.

0 favorites 0 likes
#multi-agent-reinforcement-learning

Multi-AUV Ad-hoc network-based Target Tracking: A Value Gradient Guidance Multi-Agent Diffusion Reinforcement Learning Approach

arXiv cs.LG ↗ · 2026-08-14 Cached

This paper proposes VGG-MADiffRL, a value-gradient-guided multi-agent diffusion reinforcement learning algorithm, and MDCA, a hierarchical control architecture, for cooperative target tracking in multi-AUV ad-hoc networks under constrained acoustic communication and dynamic underwater disturbances.

0 favorites 0 likes
#multi-agent-reinforcement-learning

OGR-MARL: Option-Guided Residual Multi-Agent Reinforcement Learning for Heterogeneous USV Cooperative Pursuit in Constrained Port Waterways

arXiv cs.AI ↗ · 2026-08-14 Cached

Proposes OGR-MARL, an option-guided residual multi-agent reinforcement learning framework for heterogeneous USV cooperative pursuit in constrained port waterways. The MASAC instantiation achieves a 75% capture rate and shows promising zero-shot transfer to a real map scenario.

0 favorites 0 likes
#multi-agent-reinforcement-learning

PLATO: Pointer Learner for Agent and Task Openness

arXiv cs.AI ↗ · 2026-07-29 Cached

This paper introduces PLATO, a pointer-network-based actor with a graph neural network critic for multi-agent reinforcement learning that handles both agent and task openness without retraining, evaluated in a wildfire suppression domain.

0 favorites 0 likes
#multi-agent-reinforcement-learning

Conflict Resolution under Degraded Surveillance in Air Corridors Using Multi-Agent Reinforcement Learning

arXiv cs.LG ↗ · 2026-07-24 Cached

This paper presents a deep Q-network-based multi-agent reinforcement learning framework for decentralized conflict resolution among heterogeneous small UAVs and eVTOL aircraft operating under degraded surveillance conditions, evaluating policies across 90 combinations of traffic density and separation thresholds.

0 favorites 0 likes
#multi-agent-reinforcement-learning

Feedback Attribution and Representation Geometry: Metrics for Comparing Individual and Shared Rewards in MARL

arXiv cs.LG ↗ · 2026-07-21 Cached

This paper proposes EffRank/n and D_act as low-overhead diagnostics to measure effects of reward attribution in cooperative multi-agent RL, and tests on SMACv2, finding that observation explains geometry while reward attribution mainly affects behavior.

0 favorites 0 likes
#multi-agent-reinforcement-learning

Smart charging of large fleets of Electric Vehicles: Independent Multi-Agent Reinforcement Learning approaches

arXiv cs.AI ↗ · 2026-07-01 Cached

This paper compares contextual combinatorial bandits and policy gradient algorithms for decentralized smart charging of large EV fleets, using a realistic simulation with dynamic pricing and renewable energy data.

0 favorites 0 likes
#multi-agent-reinforcement-learning

HyPOLE: Hyperproperty-Guided Multi-Agent Reinforcement Learning under Partial Observation

arXiv cs.AI ↗ · 2026-07-01 Cached

HyPOLE introduces a framework for multi-agent reinforcement learning under partial observability that uses hyperproperty-guided learning via HyperLTL temporal logic, integrated with centralized training for decentralized execution, and demonstrates improvements over baselines on SMAC, MessySMAC, and WildFire benchmarks.

0 favorites 0 likes
#multi-agent-reinforcement-learning

R2D-RL: A RoboCup 2D Soccer Environment for Multi-Agent Reinforcement Learning

arXiv cs.AI ↗ · 2026-06-18 Cached

Introduces R2D-RL, a reinforcement learning environment that connects the RoboCup 2D Soccer Simulation server to Python-based MARL workflows via shared-memory communication, supporting full-field and scenario-based training with configurable opponents and reward shaping.

0 favorites 0 likes
#multi-agent-reinforcement-learning

TRIDENT: Breaking the Hybrid-Safety-Physics Coupling for Provably Safe Multi-Agent Reinforcement Learning

arXiv cs.LG ↗ · 2026-06-18 Cached

TRIDENT is a novel multi-agent reinforcement learning framework that breaks the coupling between hybrid discrete-continuous actions, hard safety constraints, and physics-governed dynamics, achieving provably safe coordination with a convergence guarantee to a constrained Nash equilibrium and significant reductions in training-time violations.

0 favorites 0 likes
#multi-agent-reinforcement-learning

Contract-Based Compositional Shielding for Safe Multi-Agent Reinforcement Learning

arXiv cs.LG ↗ · 2026-06-15 Cached

A method for contract-based compositional shielding that ensures global safety in multi-agent reinforcement learning without centralized runtime control, using local LTL obligations and a multi-armed bandit to optimize team reward.

0 favorites 0 likes
#multi-agent-reinforcement-learning

Learn to Match: Two-Sided Matching with Temporally Extended Feedback

arXiv cs.LG ↗ · 2026-06-08 Cached

This paper introduces a framework for two-sided matching with temporally extended feedback, formulating it as a partially observable Markov game with costly screening, noisy observations, and evolving latent profiles. The authors present Learn2Match, a multi-agent reinforcement learning benchmark, and show that independent PPO outperforms bandit baselines in social welfare but incurs higher information-friction loss.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback