adaptive-computation

Tag

Cards List
#adaptive-computation

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning

arXiv cs.AI · 3d ago Cached

This paper introduces a reinforcement learning method for training a meta-reasoning policy that selects between fast reactive control and slower deliberative planning based on uncertainty in the reactive policy, achieving better balance and adaptivity in navigation tasks.

0 favorites 0 likes
#adaptive-computation

HALO: Hybrid Adaptive Latent Reasoning for Language Models

arXiv cs.CL · 2026-07-13 Cached

HALO introduces a hybrid adaptive latent refinement method for frozen language models that selectively applies second-stage refinement to a subset of tokens, achieving better performance than fixed refinement steps while using less compute.

0 favorites 0 likes
#adaptive-computation

Looped World Models

Hugging Face Daily Papers · 2026-06-16 Cached

Looped World Models introduce iterative latent state refinement through shared transformer blocks, achieving 100x parameter efficiency while adapting computational depth to prediction complexity.

0 favorites 0 likes
#adaptive-computation

AdaSR: Adaptive Streaming Reasoning with Hierarchical Relative Policy Optimization

arXiv cs.CL · 2026-06-15 Cached

Proposes AdaSR, a framework enabling reasoning models to process streaming inputs adaptively, and HRPO, a hierarchical reinforcement learning method to optimize thinking allocation for accuracy-efficiency trade-offs.

0 favorites 0 likes
#adaptive-computation

[Talk] Text Diffusion — Google DeepMind's Brendan O’Donoghue

Reddit r/LocalLLaMA · 2026-06-11 Cached

DeepMind researcher Brendan O'Donoghue provides an in-depth introduction to text diffusion models, which generate text through iterative denoising. Compared to autoregressive models, they offer lower latency but limited throughput, and demonstrate unique advantages such as self-correction and dynamic computation.

0 favorites 0 likes
#adaptive-computation

Light Interaction: Training-Free Inference Acceleration for Interactive Video World Models

Hugging Face Daily Papers · 2026-05-29 Cached

Light Interaction introduces a training-free inference acceleration framework for interactive video world models, using adaptive context management, denoising cache acceleration, and 3D block sparse attention to achieve up to 2.59x speedup while maintaining competitive visual quality.

0 favorites 0 likes
#adaptive-computation

Adaptive Computation Depth via Learned Token Routing in Transformers

arXiv cs.LG · 2026-05-08 Cached

This paper presents Token-Selective Attention (TSA), a differentiable token routing mechanism that learns to skip unnecessary computations per token in transformer layers, reducing token-layer operations by 14–23% with minimal quality loss on language modeling tasks.

0 favorites 0 likes
← Back to home

Submit Feedback