knowledge-leakage

Tag

Cards List
#knowledge-leakage

TEMPO: Temporal Enforcement via Mode-Separated Policy Optimization for Trustworthy LLM Backtesting

arXiv cs.LG · 2026-05-20

Proposes TEMPO, a policy optimization method that trains LLMs to reason exclusively from pre-cutoff information by using a two-mode reward and GRPO-based training, reducing knowledge leakage by 2–13% while improving task performance by 6–13%.

0 favorites 0 likes
#knowledge-leakage

Generating Leakage-Free Benchmarks for Robust RAG Evaluation

arXiv cs.CL · 2026-05-12 Cached

This paper introduces SeedRG, a semi-synthetic benchmark generation pipeline designed to eliminate knowledge leakage in Retrieval-Augmented Generation (RAG) evaluation by creating novel examples that preserve reasoning structures but are absent from model parametric memory.

0 favorites 0 likes
← Back to home

Submit Feedback