transformer-attention

Tag

Cards List
#transformer-attention

Mitigating LLM Over-Refusal via Dynamic Semantic Routing Calibratione

arXiv cs.CL ↗ · 2d ago Cached

The paper presents a mechanistic analysis of over-refusal in large language models and proposes Semantic Routing Calibration (SRC), a lightweight, training-free inference framework to dynamically suppress hypersensitive safety heads and mitigate over-refusal while preserving intrinsic safety.

0 favorites 0 likes
#transformer-attention

Stiefel Attention: When the Geometry of Transformer Projection Matrices Dominates Optimizer Choice---and When It Does Not

arXiv cs.LG ↗ · 2026-09-18 Cached

This paper introduces Stiefel Attention, which constrains transformer query and key projection matrices to the Stiefel manifold using Riemannian optimization, demonstrating improved performance on modular arithmetic grokking and CIFAR-10 patches.

0 favorites 0 likes
← Back to home

Submit Feedback