reasoning-optimization

Tag

Cards List
#reasoning-optimization

GrowPage: On-Demand KV Budgeting for Efficient LLM Reasoning Serving

arXiv cs.AI · 2026-09-04 Cached

GrowPage is an on-demand KV budgeting framework that dynamically manages cache capacity to enhance the efficiency of LLM reasoning serving, achieving a superior performance-throughput trade-off over existing methods.

0 favorites 0 likes
#reasoning-optimization

ParaTempo: Efficient Parallel Reasoning via Temporal Confidence

Hugging Face Daily Papers · 2026-08-17 Cached

ParaTempo is a training-free asynchronous parallel reasoning framework that uses temporal confidence to dynamically manage reasoning branches, reducing latency and token usage while maintaining accuracy in mathematical and scientific reasoning benchmarks.

0 favorites 0 likes
#reasoning-optimization

Thoughts-as-Planning: Latent World Models for Chain-of-Thoughts Optimization via Reinforcement Planning

arXiv cs.CL · 2026-05-29 Cached

Introduces Thoughts-as-Planning, a framework that models chain-of-thought optimization as sequential decision-making using latent world models and reinforcement learning, outperforming existing methods in efficiency and generalization.

0 favorites 0 likes
← Back to home

Submit Feedback