alphazero

Tag

Cards List
#alphazero

Learning to Run Power Networks: Effective AlphaZero-inspired Topological Control

arXiv cs.LG · 2026-08-17 Cached

This paper evaluates AlphaZero-inspired reinforcement learning for topological control in power networks, achieving 98.43% survivability and emphasizing the effectiveness of minimalist integration with domain heuristics.

0 favorites 0 likes
#alphazero

AlphaZero in Sparsely Rewarded Games: Limits and Auxiliary Supervision

arXiv cs.LG · 2026-07-13 Cached

This paper examines the gap between strong play and perfect play in AlphaZero for sparsely rewarded games, using Connect Four and Chomp as testbeds, and proposes an auxiliary loss (AZAL) to improve oracle consistency in optimal play.

0 favorites 0 likes
#alphazero

WallZero: Mastering the Game of WallGo with Strategic Analysis

arXiv cs.AI · 2026-06-17 Cached

This paper presents WallZero, an AlphaZero-based agent for the two-player board game WallGo, which defeats professional Go players and is used to analyze game balance and strategies.

0 favorites 0 likes
#alphazero

What to expect from AlphaZero's value predictions [D]

Reddit r/MachineLearning · 2026-05-11

The article analyzes how AlphaZero's value predictions are shaped by self-play training data and noise, questioning whether they reliably estimate win chances against opponents with different play styles despite AlphaZero's strong empirical performance.

0 favorites 0 likes
← Back to home

Submit Feedback