non-stationary-environments

Tag

Cards List
#non-stationary-environments

Closing the Feedback Loop: From Experience Extraction to Insight Governance in Verbal Reinforcement Learning

arXiv cs.AI · 2026-06-17 Cached

This paper identifies the retention-forgetting dilemma in verbal reinforcement learning for LLM agents operating in non-stationary environments, and proposes a three-layer architecture with a feedback-driven curation loop to govern insight extraction and application.

0 favorites 0 likes
#non-stationary-environments

Balancing Plasticity and Stability with Fast and Slow Successor Features

arXiv cs.LG · 2026-05-27 Cached

This paper investigates the stability-plasticity dilemma in reinforcement learning under gradual non-stationarity, finding that stabilizing successor features via synaptic consolidation across multiple timescales outperforms plasticity-focused methods.

0 favorites 0 likes
#non-stationary-environments

Turning Drift into Constraint: Robust Reasoning Alignment in Non-Stationary Environments

Hugging Face Daily Papers · 2026-05-02 Cached

The paper introduces CXR-MAX, a large-scale benchmark for evaluating reasoning alignment in non-stationary environments using X-ray data from multiple MLLMs.

0 favorites 0 likes
← Back to home

Submit Feedback