mirror-descent

Tag

Cards List
#mirror-descent

Mirror Descent Beyond Euclidean Stability: An Exponential Separation in Initialization Sensitivity

arXiv cs.LG · 2026-06-11 Cached

This paper reveals that Mirror Descent with non-quadratic regularizers can be exponentially more sensitive to initialization than Gradient Descent, even under well-conditioned settings, which has implications for reproducibility in RL and LLM post-training.

0 favorites 0 likes
#mirror-descent

Online Learning on Hidden-Convex Losses via Algorithmic Equivalence: Optimal Regret, Geometric Barrier, and Bandit Feedback

arXiv cs.LG · 2026-05-27 Cached

This paper proves that online gradient descent achieves optimal √T regret for hidden-convex losses under a Hessian compatibility condition, resolving open questions in adversarial online learning. It also extends results to one-point bandit feedback with a T^{3/4} expected regret bound.

0 favorites 0 likes
#mirror-descent

Mirror Descent-Type Algorithms for the Variational Inequality Problem with Functional Constraints

arXiv cs.LG · 2026-05-19 Cached

This paper proposes mirror descent-type algorithms for solving variational inequality problems with functional constraints, proving optimal convergence rates for problems with bounded monotone operators and Lipschitz convex constraints. A modification is introduced to improve efficiency for many constraints.

0 favorites 0 likes
#mirror-descent

Unified High-Probability Analysis of Stochastic Variance-Reduced Estimation

arXiv cs.LG · 2026-05-18 Cached

This paper presents a unified theoretical framework for stochastic variance-reduced estimation, deriving high-probability bounds via a new Freedman inequality and improving oracle complexities for constrained optimization.

0 favorites 0 likes
← Back to home

Submit Feedback