looped-language-models

Tag

Cards List
#looped-language-models

What Makes Recurrence Effective in Looped Language Models?

Hugging Face Daily Papers ↗ · 2d ago Cached

This paper systematically studies when recurrence helps in looped language models (LoopLMs), finding that extra recurrence can improve reasoning beyond the training horizon but degrade knowledge retention, and proposes channel-wise history-state injection with timestep conditioning as a more robust design for variable inference budgets.

0 favorites 0 likes
#looped-language-models

WaveFront Decoding: Parallelized Self-Speculative Decoding for Looped Language Models

Hugging Face Daily Papers ↗ · 5d ago Cached

WaveFront Decoding introduces a training-free self-speculative decoding framework for looped language models that reduces latency by concurrently batching drafting and verification, achieving up to 4.81x speedup on Huginn-3.5B.

0 favorites 0 likes
#looped-language-models

Allocating Recurrent Compute in Looped Language Models

arXiv cs.LG ↗ · 2026-08-20 Cached

This paper introduces MixerLoop, a method that allocates recurrent compute by selectively looping the mixer component in language models while applying the feed-forward network once, achieving performance improvements with reduced computational costs.

0 favorites 0 likes
#looped-language-models

Looped Language Models Improve Compositional Tool Calling

arXiv cs.AI ↗ · 2026-08-20 Cached

The paper explores how looped language models, which use iterative latent computation, improve compositional tool calling in agentic systems, showing benefits for multi-step API interactions.

0 favorites 0 likes
#looped-language-models

Looped Language Models Improve Compositional Tool Calling

Hugging Face Daily Papers ↗ · 2026-08-17 Cached

Looped language models enhance compositional tool calling by leveraging recurrent computation, improving accuracy on multi-step tasks while adaptive inference optimizes the balance between performance and compute cost. The study suggests these models are promising for reliable agentic systems.

0 favorites 0 likes
#looped-language-models

Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary

arXiv cs.LG ↗ · 2026-07-22 Cached

This paper investigates whether a frozen looped transformer can read its own computation quality (pre-answer prediction reaching AUROC 0.797) and whether external interventions can improve outcomes, finding that no tested frozen intervention produces a validated capability gain, a property termed operational proto-introspection.

0 favorites 0 likes
← Back to home

Submit Feedback