domain-agent

Tag

Cards List
#domain-agent

CLAP: Closed-Loop Training, Evaluation, and Release Control for Domain Agent Post-training

arXiv cs.AI · 2026-07-03 Cached

CLAP proposes a closed-loop method for domain agent post-training that converts noisy business data into structured SFT and preference samples, integrates reward/KL diagnosis, offline gates, and application-chain replay to decide adapter release. Experiments on five manufacturing batches show modest average gains and highlight that regression and high KL risks require an integrated data-training-evaluation-release loop rather than relying on a single score.

0 favorites 0 likes
← Back to home

Submit Feedback