rotation

Tag

Cards List
#rotation

KLQ: Training-free measured rotation quantization. Beats all training-free rotation-based quantization methods on W4A4KV4-bits. Llama 3.2 1B KLQ-quantized beats SpinQuant and gets close to ReSpinQuant without GPTQ/LDLQ rounding.

Reddit r/LocalLLaMA · 2026-08-09 Cached

KLQ is a training-free LLM quantization method that allocates bits per direction based on measured KL divergence, outperforming existing training-free rotation-based methods on W4A4KV4-bit settings for models like Llama 3.2 1B and Qwen 2.5.

0 favorites 0 likes
#rotation

Output-Aware Rotation for INT2 KV-Cache Quantization

arXiv cs.LG · 2026-08-05 Cached

Proposes OptR, an output-aware rotation method for INT2 KV-cache quantization that minimizes post-output attention error, improving QuaRot and OSCAR across models and benchmarks.

0 favorites 0 likes
#rotation

$1.3 trillion vanished Friday. AI Bubble busting, or just profit-taking?

Reddit r/ArtificialInteligence · 2026-06-08

AI stocks led a major market selloff, erasing $1.3 trillion, sparking debate on whether the AI bubble is bursting or it's just a sector rotation and profit-taking. Analysts from Goldman, Bridgewater, and BofA offer conflicting views.

0 favorites 0 likes
#rotation

Rotation revisited: Cycle decomposition in clang’s libcxx

The Old New Thing (Raymond Chen) · 2026-06-04 Cached

The article delves into the cycle decomposition algorithm used in clang's libcxx for rotation, explaining how it achieves the minimum number of swaps by computing the greatest common divisor (gcd) to determine the number of cycles.

0 favorites 0 likes
#rotation

Rotation revisited: Another unidirectional algorithm

The Old New Thing (Raymond Chen) · 2026-06-02 Cached

Raymond Chen revisits a unidirectional rotation algorithm for swapping adjacent memory blocks, explaining its recursive approach and performance characteristics.

0 favorites 0 likes
#rotation

OSCAR: Offline Spectral Covariance-Aware Rotation for 2-bit KV Cache Quantization

Hugging Face Daily Papers · 2026-05-18 Cached

OSCAR is an offline spectral covariance-aware rotation method for 2-bit KV cache quantization that aligns quantization with attention covariance structures, achieving high accuracy and efficiency for long-context LLM serving.

0 favorites 0 likes
#rotation

@GoodfireAI: Neural networks do math by rotating shapes. We found a shape-rotating calculator hidden inside an LLM – and it’s used f…

X AI KOLs Following · 2026-05-14 Cached

GoodfireAI found that neural networks perform math by rotating shapes, uncovering a shape-rotating calculator inside an LLM that is used for more than just math.

0 favorites 0 likes
← Back to home

Submit Feedback