SoL-Refiner: Speed-of-Light One-Step Refinement for High-Resolution Video
Summary
SoL-Refiner is a one-step video refiner that transforms low-resolution video outputs into 4K resolution with a single denoising step, achieving significant speed improvements and outperforming existing refineries in quality metrics.
View Cached Full Text
Cached at: 09/30/26, 08:19 AM
Paper page - SoL-Refiner: Speed-of-Light One-Step Refinement for High-Resolution Video
Source: https://huggingface.co/papers/2609.37969 Published on Sep 29
·
Submitted byhttps://huggingface.co/Owen777
Owenon Sep 30
Authors:
,
,
,
,
,
,
,
,
,
,
Abstract
High-resolutionvideogenerationisexpensive,asitscostgrowsrapidlywiththenumberofspatiotemporaltokens.Apracticalalternativefirstgeneratesalower-resolutionvideoandthenappliesarefiner,butconventionalmulti-steprefinementintroducesasecondsamplingbottleneck.WepresentSoL-Refiner,aone-stepvideorefinerthattransformslow-resolutionmodeloutputsinto4Kvideoswithasingledenoisingstep.Ourthree-stagerecipecombineshigh-resolutioncontinualtraining,reinforcementlearning(RL)post-training,andafinalone-stepdistillation.WeintroduceRefiner-Bench,avideorefinementbenchmarkconstructedfromtheoutputsofdifferentvideogenerators,anduseashared-inputprotocoltocomparerefinersatapproximately2Koutputresolution.At2K,theone-stepSoL-RefineroutperformsallexternalrefinersontheVBenchandUniPerceptaverages,whileat3840!times!2176itimprovesbothmetricsoverthethree-stepLTX-2.3Refiner.Withthecompleteaccelerationstack,SoL-Refinerachievesan8.91timesspeedupinrefinementlatencyoverthesamebaselineinour2Klatencysetting.
View arXiv pageView PDFProject pageAdd to collection
Get this paper in your agent:
hf papers read 2609\.37969
Don’t have the latest CLI?curl \-LsSf https://hf\.co/cli/install\.sh \| bash
Models citing this paper0
No model linking this paper
Cite arxiv.org/abs/2609.37969 in a model README.md to link it from this page.
Datasets citing this paper0
No dataset linking this paper
Cite arxiv.org/abs/2609.37969 in a dataset README.md to link it from this page.
Spaces citing this paper0
No Space linking this paper
Cite arxiv.org/abs/2609.37969 in a Space README.md to link it from this page.
Collections including this paper0
No Collection including this paper
Add this paper to acollectionto link it from this page.
Similar Articles
Data-Free Flow Self-Distillation for Few-Step Video Generation — open-sourced a validated implementation on MiniMax-H3 [P]
HyperFlow is an open-source 8-step LoRA that uses data-free flow self-distillation to reduce video generation steps from 49 to 8 in MiniMax-H3, achieving significant speedup while maintaining quality.
LiteFrame Scales Video LLM Efficiency (6 minute read)
LiteFrame introduces a highly efficient video encoder for Video LLMs that uses Compressed Token Distillation to enable up to 8x more frames and 35% latency reduction while maintaining accuracy, setting a new Pareto frontier for long-form video understanding.
Refiner: Robotics library from the ex-Hugging Face pre-training team
Refiner is an open-source engine from Macrodata Labs for converting raw robotics and multimodal data into high-quality datasets for model training, with local and cloud execution.
SolarFlowRefiner: Refinement-Aware Flow Matching for Surface Solar Radiation Downscaling
The paper introduces SolarFlowRefiner, a refinement-aware flow-matching framework for downscaling surface solar radiation from coarse ERA5 data to high-resolution SolarCube fields, demonstrating consistent improvements over standalone generation and post-hoc refinement methods.
HL-OutPaint: Coarse-to-Fine Video Outpainting for High-Resolution Long-Range Videos
HL-OutPaint is a coarse-to-fine video outpainting framework for high-resolution long-range videos, using global coarse guidance to enable large spatial extrapolation while maintaining spatio-temporal consistency.