research-preprint

Tag

Cards List
#research-preprint

A fixed evaluator can still become the target of an agent loop

Reddit r/artificial ↗ · 2026-08-21

The article discusses how a fixed evaluator in an agent loop can still be targeted by agent adaptation, based on the AQuA preprint, and raises questions about the isolation of validation feedback.

0 favorites 0 likes
#research-preprint

AHD Agent: Agentic Reinforcement Learning for Automatic Heuristic Design

arXiv cs.AI ↗ · 2026-05-12 Cached

This paper introduces AHD Agent, a framework using agentic reinforcement learning to enable LLMs to autonomously design heuristics for combinatorial optimization problems by dynamically interacting with the solving environment.

0 favorites 0 likes
← Back to home

Submit Feedback