OpenLongTail: Generative Scaling of Long-Tail Driving Data

Hugging Face Daily Papers Papers

Summary

OpenLongTail is an open-source generative data engine that transforms heterogeneous long-tail driving data into view-aligned multi-view assets for training robust autonomous driving policies, improving closed-loop driving robustness.

Scaling robust driving policies is fundamentally bottlenecked by the scarcity of edge cases in curated datasets. While the real world continuously captures these critical events, such long-tail events remain underutilized when collected from heterogeneous sources. Specifically, diverse but valuable in-the-wild long-tail videos lack the full view coverage required for training policy models, often missing multi-view poses or originating solely from monocular dash cameras. This modality gap prevents these ubiquitous observations from being converted into scalable training data for long-tail generalization. We introduce OpenLongTail, an open-source generative data engine for scaling autonomous driving policies under long-tail events. To transform heterogeneous data sources into view-aligned and temporally coherent multi-view assets that are useful for policy learning, we develop a pose-informed extrapolative view synthesis pipeline that generates the missing views. We further enhance cross-view consistency and the temporal alignment for the newly generated views by injecting Plücker ray geometry into the scalable generation engine. By synthesizing heterogeneous long-tail data, we observe a significant improvement in closed-loop driving robustness in handling long-tail events. By measuring the extrapolative view synthesis and pose metrics, we validate the effectiveness of OpenLongTail in visual fidelity, cross-view consistency, and ego-trajectory recovery.
Original Article
View Cached Full Text

Cached at: 07/21/26, 06:35 AM

Paper page - OpenLongTail: Generative Scaling of Long-Tail Driving Data

Source: https://huggingface.co/papers/2607.09655 Authors:

,

,

,

,

,

,

,

,

,

,

,

,

,

,

,

,

Abstract

Scalingrobustdrivingpoliciesisfundamentallybottleneckedbythescarcityofedgecasesincurateddatasets.Whiletherealworldcontinuouslycapturesthesecriticalevents,suchlong-taileventsremainunderutilizedwhencollectedfromheterogeneoussources.Specifically,diversebutvaluablein-the-wildlong-tailvideoslackthefullviewcoveragerequiredfortrainingpolicymodels,oftenmissingmulti-viewposesororiginatingsolelyfrommonoculardashcameras.Thismodalitygappreventstheseubiquitousobservationsfrombeingconvertedintoscalabletrainingdataforlong-tailgeneralization.WeintroduceOpenLongTail,anopen-sourcegenerativedataengineforscalingautonomousdrivingpoliciesunderlong-tailevents.Totransformheterogeneousdatasourcesintoview-alignedandtemporallycoherentmulti-viewassetsthatareusefulforpolicylearning,wedevelopapose-informedextrapolativeviewsynthesispipelinethatgeneratesthemissingviews.Wefurtherenhancecross-viewconsistencyandthetemporalalignmentforthenewlygeneratedviewsbyinjectingPlückerraygeometryintothescalablegenerationengine.Bysynthesizingheterogeneouslong-taildata,weobserveasignificantimprovementinclosed-loopdrivingrobustnessinhandlinglong-tailevents.Bymeasuringtheextrapolativeviewsynthesisandposemetrics,wevalidatetheeffectivenessofOpenLongTailinvisualfidelity,cross-viewconsistency,andego-trajectoryrecovery.

View arXiv pageView PDFProject pageGitHub24Add to collection

Get this paper in your agent:

hf papers read 2607\.09655

Don’t have the latest CLI?curl \-LsSf https://hf\.co/cli/install\.sh \| bash

Models citing this paper0

No model linking this paper

Cite arxiv.org/abs/2607.09655 in a model README.md to link it from this page.

Datasets citing this paper0

No dataset linking this paper

Cite arxiv.org/abs/2607.09655 in a dataset README.md to link it from this page.

Spaces citing this paper0

No Space linking this paper

Cite arxiv.org/abs/2607.09655 in a Space README.md to link it from this page.

Collections including this paper0

No Collection including this paper

Add this paper to acollectionto link it from this page.

Similar Articles

OctoLong: Mid-Training On Cross-Repository Code Contexts Enhances Long-Context Modeling

arXiv cs.AI

Introduces OctoLong, a context engineering pipeline for curating dependency-rich cross-repository code contexts, and OctoLong-Instruct, a suite of long-context open LMs trained on this data. Experiments show that replacing 12% of traditional long-context corpora with OctoLong data yields substantial gains in long-range retrieval, state tracking, repository-level code understanding, and agentic tasks.