cs.CVOct 7, 2026

PARC-Loc: Text-to-Point-Cloud Localization with Partial Assignment and Relational Consistency

Authors: Shengkai Ma, Zhenyu Hou, Weihua Cao

Abstract

Text-to-point-cloud localization estimates a position in a city-scale 3D map from descriptions of surrounding objects. Existing coarse-to-fine methods retrieve submaps using aggregate learned compatibility and then localize within a selected submap. However, repetitive or similar urban objects can inflate the embedding similarity between the query and multiple submaps, even when the instance layout within a submap violates the query description. Meanwhile, query-relevant instances often span submap boundaries, leaving the retrieved submap with incomplete contextual evidence. We term these failure modes layout-inconsistent aliasing and boundary evidence incompleteness, respectively. To address them, we propose PARC-Loc, a coarse-to-fine localization framework built on Partial Assignment with Relational Consistency (PARC). PARC jointly models hint-object compatibility and pairwise spatial relations, allowing unmatched elements while favoring assignments consistent with the queried layout. At the coarse stage, its candidate-level assessment complements neural similarity for layout-consistent submap selection. At the fine stage, the context is expanded with query-relevant instances from adjacent submaps, while PARC yields object-level matching weights that guide cross-modal attention. Extensive experiments on KITTI360Pose and CityLoc show that PARC-Loc outperforms conventional coarse-to-fine baselines. On KITTI360Pose, our method improves Top-1 localization recall at 5 m from 0.50 to 0.67, achieving a 34% relative gain over the strongest baseline.

Figures & tables

Explore similar work

CardsList
  1. PosEviLoc: Position-Conditioned Spatial Evidence for Language-Based 3D Localization

    Sep 20, 2026Tianyi Shang, Yike Shi, Zhenyu LiIndoor LocalizationPoint Clouds

  2. Reference-Induced Consensus for Selective Posed-Reference Visual Localization

    Jul 6, 2026Wonseok Kang, Jaehyun Kim, Jeongmin Lee +1Sequential Visual LocalizationIndoor Localization

  3. Efficient Sparse-to-Dense Visual Localization via Compact Gaussian Scene Representation and Accelerated Dense Pose Estimation

    May 18, 2026Zizhuo Li, Songchu Deng, Linfeng Tang +1Sequential Visual LocalizationSparse-View