3d-geometry

Tag

Cards List
#3d-geometry

See like a Robot: Robot-Centric Pointmaps for Vision-Language-Action Models

Hugging Face Daily Papers · 2026-07-13 Cached

This paper introduces robot-centric pointmaps, which encode 3D scene coordinates in the robot frame directly into image pixels, to resolve the frame mismatch between camera observations and robot action definitions in Vision-Language-Action models. The method improves Pi0.5 and SmolVLA on RoboCasa and generalizes better to unseen camera placements in real-robot experiments.

0 favorites 0 likes
#3d-geometry

World Tracing: Generative Pixel-Aligned Geometry Beyond the Visible

Hugging Face Daily Papers · 2026-06-11 Cached

World Tracing introduces a generative pixel-aligned geometry representation that predicts 3D points aligned with observed pixels while completing occluded surfaces. It uses a diffusion transformer trained with pixel-space flow matching, achieving strong performance on visible-surface reconstruction and complete geometry generation across object, scene, and dynamic benchmarks.

0 favorites 0 likes
#3d-geometry

Towards Consistent Video Geometry Estimation

Hugging Face Daily Papers · 2026-05-28 Cached

ViGeo is a transformer-based foundation model that recovers dense and consistent 3D geometry from videos using dynamic chunking attention and a completion-based data refinement framework, achieving state-of-the-art performance across multiple tasks.

0 favorites 0 likes
← Back to home

Submit Feedback