step-level-reasoning

Tag

Cards List
#step-level-reasoning

Automated Trajectory Evaluation for Mobile Agents via Step-Level Consequence Reasoning and Aggregation

arXiv cs.AI · 6d ago Cached

The paper introduces CRATE, a two-stage framework using step-level consequence reasoning to evaluate mobile agents, achieving high F1-scores on benchmarks like AndroidWorld and MobileRisk.

0 favorites 0 likes
#step-level-reasoning

Reinforcing Step-level Reasoning for Effective Self-Correction in LLMs

arXiv cs.CL · 2026-08-13 Cached

This paper introduces SFS-DPO, a reinforcement learning two-stage framework for step-level self-verification and self-correction in LLMs, with a teacher-assisted variant SFS-DPO-R. It demonstrates improvements in self-correction effectiveness across multiple LLMs with less training data than prior approaches.

0 favorites 0 likes
← Back to home

Submit Feedback