world-feedback

Tag

Cards List
#world-feedback

World Feedback for Clinical Agents: Diagnosing RL in FHIR Environments

arXiv cs.AI · 2026-07-03 Cached

This paper examines the use of reinforcement learning from world feedback for clinical protocol-execution tasks in FHIR environments, identifies structural barriers like high silent-finish ceilings and zero-gradient tasks, and introduces MedAgentBench-v3 with a lower ceiling. It shows that pure RL underperforms rule-based SFT due to these barriers, and proposes a combined SFT+RL approach.

0 favorites 0 likes
#world-feedback

Closing the Feedback Loop: From Experience Extraction to Insight Governance in Verbal Reinforcement Learning

arXiv cs.AI · 2026-06-17 Cached

This paper identifies the retention-forgetting dilemma in verbal reinforcement learning for LLM agents operating in non-stationary environments, and proposes a three-layer architecture with a feedback-driven curation loop to govern insight extraction and application.

0 favorites 0 likes
← Back to home

Submit Feedback