structured-reasoning

Tag

Cards List
#structured-reasoning

G-SHARE: A Guideline-Based Structured Reasoning Framework for Human-Factor Event Diagnosis

arXiv cs.CL · 2026-07-15 Cached

G-SHARE is a guideline-based structured reasoning framework for human-factor event diagnosis in nuclear power plants. It operationalizes a nine-step diagnostic guideline into a multi-stage pipeline with evidence extraction, stepwise reasoning, and consistency repair, outperforming one-shot LLM prompting and traditional baselines.

0 favorites 0 likes
#structured-reasoning

Test-Time Verification for Text-to-SQL via Outcome Reward Models

arXiv cs.CL · 2026-07-01 Cached

本文提出GradeSQL框架,使用结果奖励模型(ORM)进行Text-to-SQL的测试时验证,在BIRD和Spider基准上分别比基于执行的Best-of-N方法提升4.33%和2.10%。

0 favorites 0 likes
#structured-reasoning

Flow Reasoning Models: Scaling Reasoning Through Iterative Self-Refinement

arXiv cs.AI · 2026-06-30 Cached

Flow Reasoning Models (FRMs) introduce a training and test-time-scaling framework for discrete flow models on structured reasoning tasks. By using self-verification and self-conditioning, FRMs achieve nearly 100% solve rates on Sudoku and Zebra puzzles with far fewer passes than previous baselines.

0 favorites 0 likes
#structured-reasoning

Retrieval-Warmed Energy-Based Reasoning: A Five-Arm Ablation Methodology for Diffusion-as-Inference on Structured Reasoning Tasks

arXiv cs.LG · 2026-06-26 Cached

This paper presents a five-arm ablation methodology for diagnosing which component of retrieval-warmed energy-based reasoning (RW-EBR) drives performance gains, applied to structured reasoning tasks like graph reachability and Sudoku. The method separates effects of class-prior bias, stochastic warm-starting, and graph-aligned value reuse.

0 favorites 0 likes
#structured-reasoning

ScaleToT: Generalizing Structured LLM Reasoning for Billion-Scale Low-Activity User Modeling

arXiv cs.AI · 2026-06-24 Cached

ScaleToT proposes a method to generalize structured LLM reasoning for low-activity user modeling at billion scale, using tree-of-thought refinement and training a student model to reduce cost. An online A/B test in advertising deployment showed a 6.738% increase in LT30.

0 favorites 0 likes
#structured-reasoning

@mervenoyann: everyone's building simple agents meanwhile IBM is building robust enterprise agents in production, and it's open-sourc…

X AI KOLs Following · 2026-06-01 Cached

IBM released an open-source blog on Hugging Face detailing how to build robust enterprise agents with structured reasoning and tool use, going beyond basic LLMs and agents.

0 favorites 0 likes
#structured-reasoning

Pseudocode-Guided Structured Reasoning for Automating Reliable Inference in Vision-Language Models

arXiv cs.AI · 2026-05-20

Proposes the Pseudocode-guided Structured Reasoning framework (PStar) that adaptively selects structured pseudocode reasoning paths to reduce hallucinations in Vision-Language Models, achieving state-of-the-art scores on POPE and MMStar benchmarks.

0 favorites 0 likes
← Back to home

Submit Feedback