latent-reasoning

Tag

Cards List
#latent-reasoning

Hinton vs. LeCun is back: did recent reasoning models prove that LeCun was right all this time about auto-regressive LLMs?

Reddit r/singularity ↗ · 6d ago

A debate resurfaces between AI pioneers Geoffrey Hinton and Yann LeCun regarding the efficacy of autoregressive LLMs, with recent advances in reasoning models reigniting the discussion on whether transformers alone suffice for human-like reasoning.

0 favorites 0 likes
#latent-reasoning

When Steering Fails in Latent Reasoning: A Latent-to-Language Transition Gap

arXiv cs.CL ↗ · 2026-09-21 Cached

This study finds that activation steering in latent chain-of-thought reasoning is less effective than in explicit CoT, highlighting a transition gap where interventions in latent space fail to transfer to language generation.

0 favorites 0 likes
#latent-reasoning

Efficient Multimodal Generative Recommendation with Latent Narrative Reasoning

arXiv cs.CL ↗ · 2026-09-16 Cached

The paper proposes NarraLite, an efficient multimodal generative recommendation framework that uses latent narrative reasoning to improve episodic content prediction with better accuracy and efficiency.

0 favorites 0 likes
#latent-reasoning

Recurrent Looped Transformer

Hacker News Top ↗ · 2026-09-13 Cached

Recurrent Looped Transformer (RLT) is a novel architecture combining a causal encoder with a recurrent decoder to achieve latent reasoning with unbounded temporal depth, model-hardware co-design, and model-RL algorithm co-design.

0 favorites 0 likes
#latent-reasoning

Thinking that we’ll get safety by CoT traces is wishful thinking. Safety lives in the harness, not the chain of thought

Reddit r/LocalLLaMA ↗ · 2026-09-11

The article argues that relying on chain-of-thought traces for AI safety is ineffective, as they can be manipulated and do not faithfully represent model behavior, instead emphasizing the need to focus on harness control mechanisms.

0 favorites 0 likes
#latent-reasoning

Structural Process Supervision for Latent Chain-of-Thought Reasoning

arXiv cs.AI ↗ · 2026-09-11 Cached

This paper proposes Prototype-Mediated Process Supervision (PMPS) for latent chain-of-thought reasoning, achieving token compression and accuracy improvements over explicit CoT methods.

0 favorites 0 likes
#latent-reasoning

A*-Thought-V2: Efficient Latent Reasoning via Geometric Dynamics of LLM

Hugging Face Daily Papers ↗ · 2026-09-07 Cached

A*-Thought-V2 models chain-of-thought reasoning as hidden-state trajectories, using geometric dynamics to compress non-essential steps into latent tokens, improving accuracy and efficiency in LLM reasoning.

0 favorites 0 likes
#latent-reasoning

RecurTrace: Adaptive Latent Reasoning with Loop-Time Memory

arXiv cs.LG ↗ · 2026-09-04 Cached

RecurTrace introduces loop-time memory and adaptive halting to improve latent reasoning in language models, achieving higher accuracy on MathQA with optimized compute compared to fixed-loop methods.

0 favorites 0 likes
#latent-reasoning

Latent Reasoning Landscape in 2026: Mapping BDH-CQ, HRM/TRM, Coconut [D]

Reddit r/MachineLearning ↗ · 2026-09-01

The article explores latent reasoning as an alternative to chain-of-thought in AI, categorizing five families of approaches and discussing implications for AGI and interpretability.

0 favorites 0 likes
#latent-reasoning

Intelligence per dollar is the new scaling law: A tiny reasoning model breaks the existing cost-accuracy Pareto frontier on Arc-AGI 1

Reddit r/artificial ↗ · 2026-08-17

Chart Pathway's BDH-CQ, a 150M parameter reasoning model, achieves 29.5% on ARC-AGI-1 at a much lower cost per task compared to larger models like GPT-5.6 Luna, showcasing improved cost-accuracy trade-offs.

0 favorites 0 likes
#latent-reasoning

Retrieval Grounding Latent Reasoning for Dense Retrieval

arXiv cs.AI ↗ · 2026-08-17 Cached

Proposes Retrieval Grounding Latent Reasoning (RGLR), a latent reasoning framework for dense retrieval that explicitly connects intermediate latent transitions with retrieval improvements, outperforming baselines on reasoning-intensive tasks.

0 favorites 0 likes
#latent-reasoning

Think in Latent, Explain in Language: Self-Explainable Latent Reasoning

arXiv cs.CL ↗ · 2026-08-17 Cached

This paper introduces SELR, a unified framework for self-explainable latent reasoning that trains a single model to perform efficient reasoning while generating human-readable explanations, eliminating the need for external decoders.

0 favorites 0 likes
#latent-reasoning

Are Latent Reasoning Models Easily Interpretable?

Lobsters Hottest ↗ · 2026-08-15 Cached

The paper investigates the interpretability of latent reasoning models, finding that reasoning tokens are often unnecessary but can be decoded to reveal interpretable traces when needed, suggesting these models implement expected solutions.

0 favorites 0 likes
#latent-reasoning

A 150M param recurrent model scores 29.5% on ARC-AGI-1 at $0.0007 per task

Reddit r/LocalLLaMA ↗ · 2026-08-14 Cached

The article introduces BDH-CQ, a 150M parameter recurrent model that combines in-context learning with latent reasoning, achieving 29.5% on ARC-AGI-1 at a cost of $0.0007 per task, setting a new standard for cost efficiency.

0 favorites 0 likes
#latent-reasoning

Retrofitting Recurrent Depth into a Pretrained Language Model: Installation, Extrapolation, Transfer, and Retention at Two Parameter Budgets

arXiv cs.CL ↗ · 2026-08-13 Cached

This paper explores surgically retrofitting a pretrained language model (e.g., Qwen2.5-0.5B) with recurrent depth, demonstrating that the resulting model can perform deeper latent reasoning, extrapolate past supervised depth, and outperform dense models fine-tuned to reason in tokens, while also revealing catastrophic interference limits.

0 favorites 0 likes
#latent-reasoning

Did Pathway just reveal the architecture breakthrough Andrew Curran predicted? Its 150M model sets a new ARC-AGI-1 cost-efficiency frontier

Reddit r/singularity ↗ · 2026-08-11

Pathway's 150M-parameter BDH-CQ model achieves 29.5% on ARC-AGI-1 at a record-low cost of $0.0007 per task, using recurrent memory and latent reasoning instead of long token chains. The architecture may be the breakthrough Andrew Curran teased, with OpenAI researcher Lukasz Kaiser as an investor and adviser.

0 favorites 0 likes
#latent-reasoning

ENTLORE: A Graph-Grounded Benchmark for Latent Organizational Reasoning in Enterprise Question Answering

Hugging Face Daily Papers ↗ · 2026-08-11 Cached

ENTLORE is a graph-grounded benchmark framework for enterprise question answering that evaluates latent organizational reasoning, showing that even with gold document sources, many implicit relation questions remain unanswered.

0 favorites 0 likes
#latent-reasoning

Learning Latent Reasoning Traces for Scalar Reward Models End-to-End

arXiv cs.CL ↗ · 2026-08-03 Cached

Proposes LatentRM, a reward modeling framework that learns intermediate reasoning traces as discrete latent variables to explicitly maximize downstream scalar reward likelihood, improving preference modeling and policy alignment across in-distribution and OOD tasks.

0 favorites 0 likes
#latent-reasoning

GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning

Hugging Face Daily Papers ↗ · 2026-08-03 Cached

This paper introduces GradCuit, a method for test-time latent reasoning that inserts optimizable latent states at a selected Transformer layer. It achieves 64.5% average accuracy across five backbones and three reasoning benchmarks, outperforming chain-of-thought prompting and showing improved robustness and interpretability.

0 favorites 0 likes
#latent-reasoning

J-CoT: Chain-of-Thought in J-Space

arXiv cs.CL ↗ · 2026-07-27 Cached

This paper introduces J-CoT, a recurrent reasoning framework that uses vocabulary-indexed coefficients (J-thoughts) as intermediate interfaces, enabling improved reasoning performance on mathematical, scientific, coding, and path-reasoning tasks without requiring full verbalization or dense hidden state recurrence.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback