3d-reconstruction

Tag

Cards List
#3d-reconstruction

Surflo: Consistent 3D Surface Flow Model with Global State

Hugging Face Daily Papers · 2026-06-11 Cached

Surflo is a feed-forward 3D reconstruction model that compresses unposed RGB views into latent tokens and decodes consistent 3D surface points via flow matching, enabling variable-resolution output and outperforming existing methods in speed.

0 favorites 0 likes
#3d-reconstruction

BA-T: An Iterative Transformer for Two-View Bundle Adjustment

Hugging Face Daily Papers · 2026-06-02 Cached

BA-T is an iterative Transformer architecture for two-view bundle adjustment that improves 3D reconstruction accuracy and cross-view consistency using a lightweight design with only 16% of conventional decoder parameters, matching or surpassing larger models.

0 favorites 0 likes
#3d-reconstruction

Thinking in Blender: Staged Executable Inverse Graphics with Vision-Language Models

Hugging Face Daily Papers · 2026-06-01 Cached

This paper introduces SEIG, a framework that uses pretrained vision-language models to reconstruct 3D scenes from single images as editable Blender programs through progressive refinement of geometry, materials, composition, and lighting.

0 favorites 0 likes
#3d-reconstruction

SurGe: Improved Surface Geometry in Point Maps

Hugging Face Daily Papers · 2026-05-29 Cached

SurGe introduces a Neighborhood Attention Decoder and a reformulated scale-invariant gradient matching loss to improve local surface geometry accuracy in feedforward 3D reconstruction, particularly for thin structures. It achieves state-of-the-art average rank on zero-shot monocular geometry benchmarks, with better local point map and normal metrics.

0 favorites 0 likes
#3d-reconstruction

reconstructing different angles from live footage

Reddit r/singularity · 2026-05-25

4D Gaussian Splatting is a technique that converts flat 2D images into three-dimensional spatial data, enabling reconstruction of different angles from live footage.

0 favorites 0 likes
#3d-reconstruction

Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction

Hugging Face Daily Papers · 2026-05-25 Cached

Introduces GARD, a diffusion-based framework that operates in the feature space of a feed-forward 3D reconstructor to jointly recover scene geometry and high-quality imagery from degraded inputs.

0 favorites 0 likes
#3d-reconstruction

TriSplat: Simulation-Ready Feed-Forward 3D Scene Reconstruction

Hugging Face Daily Papers · 2026-05-25 Cached

TriSplat is a feed-forward 3D reconstruction network that uses oriented triangle primitives to directly generate simulation-ready meshes from single images, bypassing expensive post-processing steps. It achieves geometry-faithful reconstructions while maintaining competitive novel-view rendering quality.

0 favorites 0 likes
#3d-reconstruction

HorizonStream: Long-Horizon Attention for Streaming 3D Reconstruction

Hugging Face Daily Papers · 2026-05-22 Cached

HorizonStream introduces a long-horizon attention mechanism for streaming 3D reconstruction that explicitly models geometric propagation via an evidence influence kernel, achieving stable, scalable reconstruction with constant memory and linear time complexity, and generalizing to sequences over 10,000 frames.

0 favorites 0 likes
#3d-reconstruction

GenRecon: Bridging Generative Priors for Multi-View 3D Scene Reconstruction

Hugging Face Daily Papers · 2026-05-22 Cached

GenRecon introduces a method for 3D scene reconstruction that integrates generative 3D priors with multi-view image conditioning, achieving high-fidelity, editable mesh reconstructions of indoor environments and outperforming existing methods by 16%.

0 favorites 0 likes
#3d-reconstruction

SceneAligner: 3D-Grounded Floorplan Localization in the Wild

Hugging Face Daily Papers · 2026-05-21 Cached

Presents SceneAligner, a deep learning approach for floorplan localization that uses 3D scene reconstruction and cross-modal correspondence learning to work in real-world environments with limited data.

0 favorites 0 likes
#3d-reconstruction

UniT: Unified Geometry Learning with Group Autoregressive Transformer

Hugging Face Daily Papers · 2026-05-20 Cached

UniT is a unified feed-forward model for geometry perception using a Group Autoregressive Transformer that integrates multiple paradigms (online/offline, multi-modal, long-horizon) while maintaining metric-scale accuracy via scale-adaptive loss and queue-style KV caching. It achieves state-of-the-art performance on ten benchmarks spanning seven tasks.

0 favorites 0 likes
#3d-reconstruction

Gaussian Splat of a Strawberry

Hacker News Top · 2026-05-19 Cached

A Gaussian splat of a strawberry, created from 90 perspectives with focus stacking, using slang-splat for training and SuperSplat for viewing.

0 favorites 0 likes
#3d-reconstruction

Pompeii victim ID'd as a likely doctor

Ars Technica · 2026-05-18 Cached

Archaeologists used CT scans and 3D reconstruction to identify a Pompeii victim as a likely Roman doctor.

0 favorites 0 likes
#3d-reconstruction

EgoForce: Forearm-Guided Camera-Space 3D Hand Pose from a Monocular Egocentric Camera

Hugging Face Daily Papers · 2026-05-12 Cached

EgoForce is a monocular 3D hand reconstruction framework that uses a unified network with differentiable forearm representation, arm-hand transformers, and ray space solvers to recover absolute hand pose and position across different camera models, achieving state-of-the-art accuracy on egocentric benchmarks.

0 favorites 0 likes
#3d-reconstruction

Lite3R: A Model-Agnostic Framework for Efficient Feed-Forward 3D Reconstruction

Hugging Face Daily Papers · 2026-05-12 Cached

Lite3R is a model-agnostic framework that improves the efficiency of transformer-based 3D reconstruction using sparse linear attention and FP8-aware quantization. It reduces latency and memory usage by up to 2.4x while maintaining geometric accuracy on backbones like VGGT and DA3-Large.

0 favorites 0 likes
#3d-reconstruction

MoCam: Unified Novel View Synthesis via Structured Denoising Dynamics

Hugging Face Daily Papers · 2026-05-12 Cached

MoCam is a research paper introducing a diffusion-based framework for unified novel view synthesis that dynamically coordinates geometric and appearance priors to improve robustness against geometric errors.

0 favorites 0 likes
#3d-reconstruction

VidSplat: Gaussian Splatting Reconstruction with Geometry-Guided Video Diffusion Priors

Hugging Face Daily Papers · 2026-05-12 Cached

VidSplat is a training-free generative reconstruction framework that uses video diffusion priors to recover complete 3D scenes from sparse inputs by synthesizing novel views.

0 favorites 0 likes
#3d-reconstruction

@om_patel5: THIS GUY VIBE CODED A TRUE 3D PROPERTY TOUR TOOL WITH CLAUDE CODE not the fake 360 photo stitching that matterport does…

X AI KOLs Timeline · 2026-05-08

An individual has created a tool using Claude Code to generate true interactive 3D property tours from existing image captures, aiming to replace expensive hardware and subscriptions like Matterport.

0 favorites 0 likes
#3d-reconstruction

TT4D: A Pipeline and Dataset for Table Tennis 4D Reconstruction From Monocular Videos

Hugging Face Daily Papers · 2026-05-02 Cached

This paper introduces TT4D, a novel pipeline and large-scale dataset for reconstructing table tennis gameplay in 4D from monocular videos. It features a unique lift-first approach that estimates 3D ball trajectories and spin before time segmentation, enabling robust reconstruction even with occlusions.

0 favorites 0 likes
#3d-reconstruction

AnyRecon: Arbitrary-View 3D Reconstruction with Video Diffusion Model

Hugging Face Daily Papers · 2026-04-21 Cached

AnyRecon proposes a scalable framework for 3D reconstruction from arbitrary sparse inputs using a video diffusion model with persistent scene memory and geometry-aware conditioning.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback