agent-planning

Tag

Cards List
#agent-planning

Omni-Decision: Evidence-Ledger Planning for Omni-Modal Agents

Hugging Face Daily Papers ↗ · 2026-09-24 Cached

Omni-Decision introduces an evidence-ledger planning approach for omni-modal agents to handle noisy multimodal observations, achieving state-of-the-art accuracy on OmniGAIA and WorldSense benchmarks.

0 favorites 0 likes
#agent-planning

An agent plan needs dependencies, not just a numbered list

Reddit r/AI_Agents ↗ · 2026-09-10

The article proposes that agent plans should explicitly map task dependencies rather than relying on sequential lists, using a podcast workflow example to show how this makes waiting or repeating work more visible and efficient.

0 favorites 0 likes
#agent-planning

Repair the Amplifier, Not the Symptom: Stable World-Model Correction for Agent Rollouts

arXiv cs.AI ↗ · 2026-07-03 Cached

This paper introduces WM-SAR, a world-model correction method for agent planning that repairs causal subgraphs rather than visible symptoms, achieving better stabilization under token budgets compared to standard LLM correctors.

0 favorites 0 likes
#agent-planning

CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing

Hugging Face Daily Papers ↗ · 2026-05-06 Cached

The paper introduces CreativityBench, a benchmark for evaluating large language models' ability to creatively repurpose tools based on affordance reasoning. It highlights that current models struggle with creative problem-solving despite strong general reasoning capabilities.

0 favorites 0 likes
← Back to home

Submit Feedback