multi-timescale

Tag

Cards List
#multi-timescale

Learning from Environmental Feedback: Credit Assignment across Multiple Timescales for Agentic Reinforcement Learning

arXiv cs.LG · 2026-08-11 Cached

This paper introduces EFCA, a multi-timescale credit assignment method for agentic reinforcement learning that uses short-term feedback and medium-term state-history signals from environment interaction to improve task success and quality on ALFWorld and WebShop.

0 favorites 0 likes
#multi-timescale

AutoPersonas: A Multi-Timescale Loop Engine for Open-Ended Persona Evolution

arXiv cs.AI · 2026-07-10 Cached

This paper introduces AutoPersonas, an architecture for long-term persona agents that prevents 'self-locking' by separating divergence from evidence-governed absorption, demonstrating reduced repetition and improved identity continuity in simulated environments.

0 favorites 0 likes
#multi-timescale

Performance-Driven Environment Abstraction with Multi-Timescale Learning

arXiv cs.LG · 2026-06-17 Cached

This paper proposes a performance-driven state abstraction method for reinforcement learning that directly optimizes decision quality, using a multi-timescale framework to jointly adapt the policy and a tree-structured abstraction. The algorithm refines or aggregates state space based on Q-value discrepancies, achieving better sample efficiency and faster replanning than baselines.

0 favorites 0 likes
#multi-timescale

Representation over Routing: Overcoming Surrogate Hacking in Multi-Timescale PPO

Hugging Face Daily Papers · 2026-05-21 Cached

This paper identifies surrogate hacking and temporal uncertainty as failure modes in multi-timescale RL, and proposes a Target Decoupling architecture that removes routing from the actor, using the critic for auxiliary representation learning. The method eliminates policy collapse on the LunarLander-v2 benchmark and stably surpasses the 'Environment Solved' threshold without hyperparameter hacking.

0 favorites 0 likes
← Back to home

Submit Feedback