sequential-decision-making

Tag

Cards List
#sequential-decision-making

Let it Cook: Learning to Wait in Sequential Decision Making

arXiv cs.LG · yesterday Cached

This paper introduces a reinforcement learning approach for training agents to wait strategically in sequential decision-making tasks, balancing task performance with resource conservation. Experiments show significant waiting behaviors across household and continuous-state environments.

0 favorites 0 likes
#sequential-decision-making

Rushes: A Human Preference Dataset for Pluralistic Alignment

arXiv cs.CL · 2026-07-24 Cached

Introduces Rushes, a large-scale dataset of human engagement preferences in AI-generated branching narratives, revealing that current LLMs like GPT-5 fail to outperform simple baselines in predicting user choices, highlighting the need for personalized alignment.

0 favorites 0 likes
#sequential-decision-making

Can Induced Emotion Bias LLM Behaviors in Sequential Decision Making?

arXiv cs.CL · 2026-07-15 Cached

This paper investigates whether induced emotions can bias the sequential decision-making of LLMs using the Iowa Gambling Task as a testbed. The authors find that while emotional induction does not significantly affect average decision dynamics, anger can reduce penalty sensitivity and early-stage exploration.

0 favorites 0 likes
#sequential-decision-making

Stochastic Linear Bandits with Partially Observed Actions

arXiv cs.LG · 2026-07-13 Cached

This paper studies stochastic linear bandits where the agent only observes a random subset of action coordinates, proving that sublinear regret is possible when actions have low intrinsic dimension, and proposes the TOFU-POV algorithm with theoretical guarantees.

0 favorites 0 likes
#sequential-decision-making

Internalizing the Future: A Unified Agentic Training Paradigm for World Model Planning

arXiv cs.AI · 2026-06-29 Cached

Introduces a three-stage training paradigm to internalize world model planning in LLM agents, enabling future-aware decision-making. Outperforms baselines on search and mathematical reasoning tasks.

0 favorites 0 likes
#sequential-decision-making

Beyond Next-Observation Prediction: Agent-Authored World Modeling for Sequential Decision Making

arXiv cs.CL · 2026-06-25 Cached

This paper introduces Agent-Authored World Modeling (AAWM), a training procedure that constructs world-model supervision based on the policy's own decision needs rather than next-observation prediction, aligning the learning objective with the dynamics required for effective decision-making.

0 favorites 0 likes
#sequential-decision-making

Agentick: A Unified Benchmark for General Sequential Decision-Making Agents

arXiv cs.AI · 2026-05-11 Cached

This paper introduces Agentick, a unified benchmark for evaluating general sequential decision-making agents across RL, LLM, and VLM paradigms. It provides 37 procedurally generated tasks and reveals that no single approach currently dominates, highlighting significant room for improvement in agent autonomy.

0 favorites 0 likes
#sequential-decision-making

PRISM: Perception Reasoning Interleaved for Sequential Decision Making

arXiv cs.AI · 2026-05-08 Cached

This paper introduces PRISM, a framework that integrates Vision-Language Models and Large Language Models through a dynamic question-answering pipeline to improve sequential decision-making in embodied AI tasks.

0 favorites 0 likes
← Back to home

Submit Feedback