low-rank-attention

Tag

Cards List
#low-rank-attention

FLARE++: Low-rank attention with dynamic attention routing

arXiv cs.LG · 2026-08-13 Cached

FLARE++ is a low-rank attention architecture that replaces static learned queries with input-conditioned dynamic routing, improving on FLARE across PDE surrogate benchmarks and Long Range Arena while preserving linear complexity.

0 favorites 0 likes
#low-rank-attention

Low-Rank Attention Residuals

arXiv cs.LG · 2026-07-14 Cached

This paper introduces Low-Rank Attention Residuals (LR-AttnRes) for LLMs, which decouple routing from representation by using low-dimensional keys for depth-wise attention, improving performance while reducing FLOPs.

0 favorites 0 likes
#low-rank-attention

Hierarchical Attention via Domain Decomposition

arXiv cs.LG · 2026-06-18 Cached

Proposes a hierarchical attention mechanism using overlapping Schwarz domain decomposition to replace dense global low-rank attention with a two-level additive structure of local and coarse blocks, showing faster training and better accuracy with fewer parameters.

0 favorites 0 likes
← Back to home

Submit Feedback