cs.CVOct 7, 2026

LVSPM: Long Sequence View Synthesis and Pose Estimation Model

Authors: Xi Chen, Yachi Zhang, Linghao Chen, Minghua Liu, Hao Su, Zexiang Xu, Xiaoshuai Zhang

Organizations: UC San Diego · Sudo AI GmbH

Abstract

We present LVSPM, a generalizable model that jointly estimates camera poses and synthesizes novel views from uncalibrated image collections. Trained with only RGB images and pose supervision, LVSPM avoids dense 3D ground truth and employs test-time training (TTT) layers to scale seamlessly to hundreds of input views. On RealEstate10k, Co3Dv2, and DL3DV, LVSPM surpasses VGGT in pose estimation across 16-256 views, with especially large margins at strict thresholds. For novel view synthesis under a practical protocol where more views cover larger scenes, LVSPM achieves state-of-the-art pose-free quality---surpassing even pose-dependent models in PSNR---and still maintains high quality as scene scale grows, while baselines collapse. The code is available at https://burningdust21.github.io/Projects/LVSPM .

Explore similar work

CardsList
  1. SPFSplatV2: Efficient Self-Supervised Pose-Free 3D Gaussian Splatting from Sparse Views

    Sep 21, 2025Ranran Huang, Krystian Mikolajczyk3D Gaussian Splatting3D Reconstruction

  2. DVSM: Decoder-only View Synthesis Model Done Right

    May 28, 2026Cheng Sun, Jaesung Choe, Min-Hung Chen +2Novel View Synthesis

  3. PoseCompass: Intelligent Synthetic Pose Selection for Visual Localization

    May 12, 2026Yanan Zhou, Zhaoyan Qian, Yanli Li +3Visual Place RecognitionCamera Pose Estimation