3d-reconstruction

Tag

Cards List
#3d-reconstruction

@ZiyunClaudeWang: What if every meter of robot motion were optimized for reconstruction? Excited to share TRACE, our new work on active 3…

X AI KOLs Timeline · 5d ago Cached

TRACE introduces a novel approach to active 3D reconstruction by optimizing full sensor trajectories for ergodic coverage of scene information, outperforming next-best-view baselines with a 1.5 dB PSNR improvement.

0 favorites 0 likes
#3d-reconstruction

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models

Hugging Face Daily Papers · 2026-08-05 Cached

Introduces UniWorld-View, a unified framework for large-baseline novel view synthesis from monocular inputs, integrating occlusion-aware point cloud rendering with video diffusion models for precise camera control and geometric consistency.

0 favorites 0 likes
#3d-reconstruction

Articulated Object Reconstruction from Rest-State Observation

Hugging Face Daily Papers · 2026-07-30 Cached

This paper introduces a rest-state framework that reconstructs articulated objects from a single closed configuration, using explicit meshes, vision-language outputs, and video diffusion models to generate and validate articulation hypotheses without observed motion.

0 favorites 0 likes
#3d-reconstruction

@vincieye: Monocular video to 3D in real time? MAGiSt3R uses multiple agents & a merging model (MAGMA) to reconstruct scenes from …

X AI KOLs Timeline · 2026-07-22 Cached

MAGiSt3R is a multi-agent feed-forward framework that achieves real-time 3D reconstruction from monocular RGB videos at 10 FPS, using a merging model (MAGMA) to combine local point maps and pose graph optimization to reduce drift.

0 favorites 0 likes
#3d-reconstruction

@FinanceYF5: A Chinese open-source model can reconstruct 3D scenes in real-time from any video. Just one camera, no LiDAR needed, handles over 10,000 frames without crashing, and runs at 20fps on a single GPU. Its benchmark scores surpass traditional optimization-based methods, tested on drone aerial photography, dashcams, and indoor tours. And it's fully open-source.

X AI KOLs Timeline · 2026-07-20 Cached

A Chinese open-source model reconstructs 3D scenes in real-time from a single camera video, no LiDAR required, achieves 20fps on a single GPU with performance superior to traditional optimization methods. Fully open-source.

0 favorites 0 likes
#3d-reconstruction

Designing a 4D Digital Archive for Ikebana

Hacker News Top · 2026-07-19 Cached

This paper proposes a 4D digital viewing system for archiving the creation process of Ikebana using synchronized cameras, eye tracking, and 3D Gaussian Splatting, enabling interactive viewing from arbitrary angles.

0 favorites 0 likes
#3d-reconstruction

@huchris007: In 1933, Liang Sicheng surveyed the Yingxian Wooden Pagoda. He wrote a letter to Lin Huiyin at the foot of the pagoda, saying it was so magnificent that he wished she could see it right then. I also visited it last March and was truly awed. Last night, I used Kimi K3 to recreate a digital version of the Yingxian Wooden Pagoda. 10,611 components, four lighting scenes: morning, noon, dusk, and night. K3 did it all in one go. Awesome!

X AI KOLs Timeline · 2026-07-19 Cached

The user used the Kimi K3 model to reconstruct a digital version of the Yingxian Wooden Pagoda, containing 10,611 components and four lighting scenes (morning, noon, dusk, night), demonstrating the model's 3D reconstruction capability.

0 favorites 0 likes
#3d-reconstruction

@thesupermanmx: China open-sourced a model that reconstructs any scene in 3D from a regular video, in real-time. one camera. no LiDAR. …

X AI KOLs Timeline · 2026-07-16

China open-sourced a real-time 3D scene reconstruction model that works from a single regular video without LiDAR, achieving 20 FPS on a single GPU and maintaining stability over 10,000+ frames.

0 favorites 0 likes
#3d-reconstruction

Repairing Shape-Prior Shortcuts in Long-Range Single-Shot Fringe Projection Profilometry

arXiv cs.LG · 2026-07-15 Cached

Introduces PhiCalNet, a neural network for single-shot fringe projection profilometry that avoids shape-prior shortcuts by outputting wrapped phase and using a fixed calibration layer, achieving 3.3x lower error than baseline UNet on a synthetic benchmark.

0 favorites 0 likes
#3d-reconstruction

To a depth camera, a glass wall is basically empty space. This model fills it back in.

Reddit r/singularity · 2026-07-07

A model is proposed to fill in missing depth data from depth cameras when encountering transparent surfaces like glass walls, addressing a common sensor limitation.

0 favorites 0 likes
#3d-reconstruction

PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space

Hugging Face Daily Papers · 2026-07-06 Cached

PixWorld presents a unified pixel-space diffusion approach for 3D scene reconstruction and generation, overcoming limitations of latent-space methods by using direct image-level supervision and geometry-aware feature alignment. The method outperforms prior generation methods and matches state-of-the-art reconstruction methods.

0 favorites 0 likes
#3d-reconstruction

@FinanceYF5: Someone connected Blender to Fable 5, and in about 20 minutes, it reconstructed the entire building complex of New York City. It didn't just randomly draw; first, it grabbed building data from public data sources, then modeled according to the data — theoretically, the entire model is to scale. The poster said this approach of 'checking data before acting' is smarter than Opus 4.…

X AI KOLs Following · 2026-07-04 Cached

Someone connected Blender to Fable 5, and in just 20 minutes, it reconstructed the entire New York City building complex based on public building data, with the model to scale. The poster believes this 'check data before acting' approach is smarter than Opus 4.8, reflecting that AI is beginning to have the ability to 'do its homework'.

0 favorites 0 likes
#3d-reconstruction

PointDiT: Pixel-Space Diffusion for Monocular Geometry Estimation

Hugging Face Daily Papers · 2026-07-02 Cached

PointDiT presents a minimalist pixel-space diffusion transformer using a plain ViT architecture for monocular geometry estimation, outperforming complex latent-based models while maintaining simplicity and robustness in ambiguous regions.

0 favorites 0 likes
#3d-reconstruction

Scenes as Objects, Not Primitives: Instance-Structured 3D Tokenization from Unposed Views

Hugging Face Daily Papers · 2026-06-28 Cached

This paper proposes a feed-forward framework that decomposes 3D scenes into instance-structured token groups from unposed multi-view images, enabling direct object-level reconstruction, segmentation, and manipulation without 3D annotations.

0 favorites 0 likes
#3d-reconstruction

Lift4D: Harmonizing Single-View 3D Estimation for 4D Reconstruction In-the-Wild

Hacker News Top · 2026-06-23 Cached

Lift4D is a test-time optimization framework that reconstructs complete 4D geometry, appearance, and deformation of dynamic objects from a single monocular in-the-wild video, improving over prior methods on challenging sequences with occlusions and non-rigid motion.

0 favorites 0 likes
#3d-reconstruction

The cube, the epicycles and the human face

Lobsters Hottest · 2026-06-20 Cached

The article describes a project that uses Fourier series and epicycles to reconstruct a human face on the three faces of a cube, demonstrating how sinusoids can generate complex shapes.

0 favorites 0 likes
#3d-reconstruction

GeneralVLA-2: Geometry-Aware Reconstruction and Governed Memory for Robot Planning

Hugging Face Daily Papers · 2026-06-16 Cached

GeneralVLA-2 introduces GeoFuse-MV3D for improved 3D reconstruction and a governed KnowledgeBank for better memory management in robotic manipulation tasks, achieving performance gains on several benchmarks.

0 favorites 0 likes
#3d-reconstruction

SpatialAvatar-0: High-Quality 4D Head Avatar with Multi-Stage Reconstruction

Hugging Face Daily Papers · 2026-06-14 Cached

SpatialAvatar-0 introduces a multi-stage reconstruction method for high-quality 4D head avatars using a shared FLAME-mesh-bound Gaussian representation, achieving superior performance across benchmarks with reduced iterations.

0 favorites 0 likes
#3d-reconstruction

World Tracing: Generative Pixel-Aligned Geometry Beyond the Visible

Hugging Face Daily Papers · 2026-06-11 Cached

World Tracing introduces a generative pixel-aligned geometry representation that predicts 3D points aligned with observed pixels while completing occluded surfaces. It uses a diffusion transformer trained with pixel-space flow matching, achieving strong performance on visible-surface reconstruction and complete geometry generation across object, scene, and dynamic benchmarks.

0 favorites 0 likes
#3d-reconstruction

Surflo: Consistent 3D Surface Flow Model with Global State

Hugging Face Daily Papers · 2026-06-11 Cached

Surflo is a feed-forward 3D reconstruction model that compresses unposed RGB views into latent tokens and decodes consistent 3D surface points via flow matching, enabling variable-resolution output and outperforming existing methods in speed.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback