lookahead-bias

Tag

Cards List
#lookahead-bias

Scaling Point-in-Time Language Models

arXiv cs.CL ↗ · 2026-07-15 Cached

This paper demonstrates that scaling point-in-time language models—trained exclusively on text available up to each calendar date—can substantially narrow the performance gap with unrestricted models, enabling valid backtests and causal inference in finance and social sciences. The authors train decoder-only transformers up to 4B parameters on 1 trillion chronologically filtered tokens and release the full pipeline.

0 favorites 0 likes
← Back to home

Submit Feedback