To a depth camera, a glass wall is basically empty space. This model fills it back in.
Summary
A model is proposed to fill in missing depth data from depth cameras when encountering transparent surfaces like glass walls, addressing a common sensor limitation.
Similar Articles
@heyshrutimishra: Nobody talks about this but every robot on the market is blind to glass. Put a mirror in front of it. A glass bottle. I…
LingBot-Depth 2.0, trained on 150M samples, solves the longstanding problem of robots being blind to glass and transparent objects, achieving top performance on 12/16 depth benchmarks and halving depth error. Ant Group used it to significantly improve their robots' perception.
One Scene, Two Depths: Probing Geometric Ambiguity in Monocular Foundation Models
Introduces MultiDepth-3k, a benchmark to evaluate depth-layer preferences in monocular depth foundation models, and shows Laplacian Visual Prompting can alter reported depth layers, suggesting complementary geometric hypotheses exist across models.
chenxwh/depth-anything-v2
Depth Anything V2 is a monocular depth estimation model that significantly outperforms V1 in fine-grained details and robustness, offering faster inference and higher accuracy than SD-based models. It is available on Replicate under varying licenses.
@reczko_konrad: Depth-aware light injection in TypeGPU I got a 448x448 monocular depth model down to ~8 ms on my M4 Pro across ~250 dis…
A developer achieved real-time monocular depth estimation by optimizing a 448x448 model to run at ~8 ms on an M4 Pro using TypeGPU, allowing seamless GPU integration for depth-aware lighting.
Unified Panoramic Geometry Estimation via Multi-View Foundation Models
PaGeR adapts the multi-view perspective foundation model Depth Anything 3 to predict scale-invariant and metric depth, surface normals, and sky segmentation from a single equirectangular image, using a fixed cubemap representation that keeps VRAM and runtime constant. The paper also releases the ZüriPano and PanoInfinigen datasets.