linear-time

Tag

Cards List
#linear-time

ALPHABET: A Laplace-Pole History Aggregator with Banked Exponential Transport

arXiv cs.LG · 21h ago Cached

ALPHABET is a compact linear-time sequence model that compresses temporal history into stable complex pole modes, achieving competitive performance with fewer parameters and faster inference.

0 favorites 0 likes
#linear-time

An Update on Matrix Recurrent Units, an Attention Alternative [R]

Reddit r/MachineLearning · 2026-06-21

An update on Matrix Recurrent Units (MRU), a linear-time attention alternative. The author explores methods to stabilize training, finding that orthogonal matrices underperform while LDU factorization works best, and shows MRU underperforms transformers on larger datasets like TinyStories.

0 favorites 0 likes
#linear-time

Gaussian Mixture Attention: Linear-Time Sequence Mixing via Probabilistic Latent Routing

arXiv cs.LG · 2026-06-18 Cached

This paper introduces Gaussian Mixture Attention (GMA), a probabilistic attention mechanism that replaces explicit pairwise query-key comparisons with routing through learned Gaussian mixture components, achieving linear-time complexity in sequence length. Experiments show competitive performance on long-context tasks with fixed-K linear memory scaling.

0 favorites 0 likes
#linear-time

Linear Scaling Video VLMs for Long Video Understanding

Hugging Face Daily Papers · 2026-05-29 Cached

StateKV is an inference-time method that enables linear-time video prefill for long-video vision-language models by carrying cross-frame context in a fixed-capacity recurrent state, maintaining accuracy close to full self-attention without fine-tuning.

0 favorites 0 likes
← Back to home

Submit Feedback