Tag
This paper introduces TriLayer, a large-scale video dataset with foreground-background-composite triplets, and DBL-Diffusion, a dual-branch diffusion framework for explicit layered video representation, enabling high-fidelity object insertion and layer decomposition.
This paper introduces CoIn, a novel framework for 3D scene inpainting that bridges 2D diffusion models and 3D Gaussian Splatting via a multi-stage consistency pipeline, enabling both object removal and insertion with flexible masks.
This paper introduces DIRECT, a framework for pose-controllable 3D-aware object insertion that decomposes conditions into appearance, geometry, and context guidance to achieve high-fidelity compositing with explicit 3D pose control.