cs.CVOct 8, 2026

Pose-Free Feed-Forward 3D Inpainting via Learnable Mask Attention and Support Token Refinement

Authors: Jingyi Pan, Dan Xu, Qiong Luo

Organizations: The Hong Kong University of Science and Technology (Guangzhou) · The Hong Kong University of Science and Technology

Abstract

3D scene inpainting aims to recover missing or occluded regions in edited 3D scenes, while ensuring geometric and textural consistency. Existing approaches, however, typically require accurately calibrated camera poses, which restricts their applicability in casual, in-the-wild scenarios and introduces additional preprocessing overhead. To overcome this limitation, we present FreeInpaint, a novel feed-forward framework that generates complete and 3D-consistent scenes directly from unposed multi-view images with masked regions. At its core, FreeInpaint extends a 3D foundation model to propagate masked regions from a reference view to other unposed views, bridging 3D reconstruction and scene inpainting while preserving the model's native ability to recover camera poses and scene geometry. Our method addresses two key challenges in adapting feed-forward 3D foundation models to masked inputs. First, masked regions can corrupt cross-view correspondence reasoning, degrading pose estimation and geometry recovery. To address this, we introduce a Learnable Mask Attention mechanism that preserves the spatial anchoring of reliable observations while allowing masked regions to progressively absorb useful context in deeper layers. Second, under severe occlusions, a single forward pass often lacks sufficient appearance evidence for high-fidelity completion. Therefore, we propose a Support Token Refinement strategy, which injects diffusion-generated support evidence as confidence-weighted auxiliary tokens to refine under-observed regions while preserving the original spatial anchor. Extensive experiments across diverse datasets demonstrate that FreeInpaint achieves superior inpainting quality, eliminating the reliance on pre-computed camera poses while keeping a fast inference speed. The project page is https://rorisis.github.io/FreeInpaint/.

Figures & tables

Appendix figures & tables13 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. CoIn: Comprehensive 2D-3D Inpainting with Gaussian Splatting Guidance

    Jun 25, 2026Hana Kim, Minje Kim, Tae-Kyun Kim3D Gaussian Splatting3D Scene Editing

  2. WINGS: Reference-Free Gaussian Splatting Inpainting with 3D-Native Generative Priors

    Sep 29, 2026Noé Lallouet, Michael Fischer, Elie Michel3D Gaussian Splatting3D Scene Editing

  3. Generalizable Sparse-View 3D Reconstruction from Unconstrained Images

    Apr 30, 2026Vinayak Gupta, Chih-Hao Lin, Shenlong Wang +23D Gaussian Splatting3D Reconstruction