compute-optimal-training

Tag

Cards List
#compute-optimal-training

Scaling Laws for Behavioral Foundation Models over User Event Sequences

arXiv cs.LG ↗ · 2026-06-05 Cached

This paper studies scaling laws for behavioral foundation models trained on sequences of user actions, finding that a small event embedder is compute-optimal and that the evaluation metric itself influences the optimal compute allocation.

0 favorites 0 likes
← Back to home

Submit Feedback