minimax-regret

Tag

Cards List
#minimax-regret

Continuity-Free Near-Minimax Leading-Order Regret for CVaR-UCBVI

arXiv cs.LG ↗ · 2026-09-01 Cached

The paper shows that the Bernstein CVaR-UCBVI algorithm achieves a near-minimax leading-order regret bound for CVaR reinforcement learning without continuity assumptions on return laws.

0 favorites 0 likes
#minimax-regret

Information Routing across Batch Boundaries: Memory--Batch Tradeoffs in Lipschitz Bandits

arXiv cs.LG ↗ · 2026-08-11 Cached

This paper studies the joint effect of memory width and batch depth in stochastic Lipschitz bandits, characterizing the minimax pseudo-regret tradeoff up to logarithmic factors and showing that state width and update depth are not interchangeable.

0 favorites 0 likes
← Back to home

Submit Feedback