low-rank-bias

Tag

Cards List
#low-rank-bias

The Loss Does Not See the Basis, but Adam Does

Hugging Face Daily Papers · 2026-08-05 Cached

This paper investigates why Adam does not exhibit gradient descent's implicit low-rank bias in factored models, showing that coordinate-wise preconditioning breaks the relevant symmetry, while shared-scalar methods like Muon and Shampoo preserve it.

0 favorites 0 likes
#low-rank-bias

The Implicit Bias of Depth: From Neural Collapse to Softmax Codes

arXiv cs.LG · 2026-05-25 Cached

This paper studies how depth alone induces an implicit low-rank bias in deep unconstrained feature models trained without regularization, shifting the optimal solution from neural collapse to softmax codes, and provides the first asymptotic and dynamic characterization of this bias under gradient descent with cross-entropy loss.

0 favorites 0 likes
← Back to home

Submit Feedback