training-objective

Tag

Cards List
#training-objective

Principled Thoughts for Latent Recursive LLM Systems

Hugging Face Daily Papers ↗ · 5d ago Cached

The paper presents REST, a novel training objective for latent recursive LLM systems that enhances accuracy by up to 7.5 percentage points across benchmarks by incorporating properties like causality and minimality into differentiable losses.

0 favorites 0 likes
#training-objective

Improving LLMs via Validator-to-Generator Alignment

arXiv cs.CL ↗ · 2026-07-07 Cached

A new method, FLORA (Frequency-corrected Learning of Ordered Rank Alignment), improves LLMs by aligning generator and validator modes using a principled frequency correction. Experiments show substantial gains in G-V consistency and generator performance on benchmarks like IFEval and HumanEval.

0 favorites 0 likes
← Back to home

Submit Feedback