Ego-OSCAR: Egocentric Open source Stereo CAptuRe System
Summary
Ego-OSCAR presents an open-source, low-cost head-mounted stereo-inertial capture device for egocentric data collection, releasing hardware designs, software, and about 550 hours of annotated stereo video with IMU data.
View Cached Full Text
Cached at: 08/11/26, 10:20 AM
Paper page - Ego-OSCAR: Egocentric Open source Stereo CAptuRe System
Source: https://huggingface.co/papers/2608.08285
Abstract
WepresentEgo-OSCAR,anopen-hardware,low-cost,head-mountedstereo-inertialcapturedeviceforegocentricdatacollectioninthewild.EgoOSCARpairsahardware-synchronizedglobal-shutterstereocamerawitha6-axisIMU,anembeddedLinuxSBCforon-devicevideoencoding,andarealtimemicrocontrollerforuserfeedbackandwatchdogfunctions.ThecompletebillofmaterialsisunderUSD200perunit,usingonlycommerciallyavailablecomponentsand3D-printedparts.Alongsidethedevice,wereleaseacompletesoftwarestack(hardware-acceleratedrecordingpipeline,IMUsamplingdaemon,time-synchronizationtooling,andwatchdogfirmware)androughly550hoursofegocentricstereovideopercamerawithsynchronizedIMU,collectedbyadistributedcontributornetworkacrosseverydayindoorenvironments.Thereleaseisannotatedratherthanraw:free-formactioncaptionscoveressentiallytheentirerecordedtimelinewithanopenvocabulary,andper-frame3Dhandreconstructionsshipalongsideper-sessionstereocalibration.Ego-OSCARdoesnotaimtomatchtheper-unitfidelityofresearch-gradesystemssuchasProjectAria;itaimstobethecheapestdefensiblesubstrateforcrowdsourcedegocentriccapture,andtolowertheactivationenergyforanyteamthatwantstocollectegocentricdataatscale.Allhardwaredesigns,software,andthedatasetareopen-sourced
View arXiv pageView PDFAdd to collection
Get this paper in your agent:
hf papers read 2608\.08285
Don’t have the latest CLI?curl \-LsSf https://hf\.co/cli/install\.sh \| bash
Models citing this paper0
No model linking this paper
Cite arxiv.org/abs/2608.08285 in a model README.md to link it from this page.
Datasets citing this paper1
#### fpvlabs/stereo-550 Updatedabout 2 hours ago • 184
Spaces citing this paper0
No Space linking this paper
Cite arxiv.org/abs/2608.08285 in a Space README.md to link it from this page.
Collections including this paper0
No Collection including this paper
Add this paper to acollectionto link it from this page.
Similar Articles
EgoForce: Forearm-Guided Camera-Space 3D Hand Pose from a Monocular Egocentric Camera
EgoForce is a monocular 3D hand reconstruction framework that uses a unified network with differentiable forearm representation, arm-hand transformers, and ray space solvers to recover absolute hand pose and position across different camera models, achieving state-of-the-art accuracy on egocentric benchmarks.
EgoSteer: A Full-Stack System Towards Steerable Dexterous Manipulation from Egocentric Videos
EgoSteer presents a full-stack system that pre-trains a vision-language-action model from egocentric human videos for steerable dexterous manipulation, enabling robust generalization across 40+ diverse tasks with 75%+ success on complex long-horizon tasks.
MobileEgo Anywhere: Open Infrastructure for long horizon egocentric data on commodity hardware
MobileEgo Anywhere is a mobile-based framework for collecting long-duration egocentric robot data using smartphone sensors, enabling large-scale training of vision-language-action models by lowering hardware barriers.
@macrodata_labs: Everyone is betting on Egocentric data to scale robotics But turning that footage into training data requires recoverin…
Macrodata Labs releases a research blog on scaling robotics with egocentric video data by recovering 3D hand motion signals using only open-source models.
Ego2Robot: Scalable Robot Data Synthesis from Egocentric Human Data
Ego2Robot is a scalable pipeline that converts egocentric human manipulation videos into robot training data via action retargeting and visual synthesis, producing 18,561 hours of data across 15 robot morphologies. Experiments show that joint pretraining on this synthesized data improves out-of-distribution generalization for vision-language-action models, including on real-robot deployment.