cs.ROSep 28, 2026

ExcavaTwin: Training-Free Geometry-Guided Semantic Elevation Mapping for Autonomous Excavation

Authors: Yu Deng, Lingshan Zeng, Tong Hu, Rushi Dai

Organizations: The Smart Manufacturing Thrust, The Hong Kong University of Science and Technology (Guangzhou) · Department of Electronic and Electrical Engineering, Southern University of Science and Technology. · Capstone Technology Co., Ltd., Shenzhen, China.

Abstract

Autonomous excavation requires a spatial representation that jointly captures terrain geometry and task-relevant semantics. Existing excavation mapping is largely elevation-centric, while generic semantic models remain unstable in unstructured outdoor scenes. We present ExcavaTwin, a pure-vision geometry-guided semantic elevation mapping framework without excavation-specific training. Given multi-view RGB images, the framework: 1) reconstructs scene geometry and semantic observations using frozen vision models; 2) derives terrain and non-terrain geometric support; 3) performs geometry-constrained multi-view semantic fusion to suppress implausible predictions and recover incomplete observations; and 4) projects the fused state into a task-oriented semantic elevation map. Experiments on public datasets and real excavation scenes demonstrate reliable geometric and semantic perception. In real excavation, the system achieved an average update interval of approximately 1.4 s and a mean elevation error of 12.74cm in dynamically modified regions. Larger errors mainly occur during rapid terrain changes and transient visual disturbances caused by machine motion.

Figures & tables

Explore similar work

CardsList
  1. Semantic-Aware Guided Drone Exploration for Language-Conditioned 3D Indoor Mapping

    May 22, 2026Nitin Vegesna, Avideh ZakhorAerial RoboticsUAV Navigation

  2. RoboAtlas: Contextual Active SLAM

    Jun 24, 2026Alexander Schperberg, Shivam K. Panda, Abraham P. Vinod +2Simultaneous Localization and MappingRobot Navigation

  3. ActiveLang: Active Open-Vocabulary 3D Mapping with Semantic-Uncertainty-Guided Exploration

    Oct 7, 2026Liyan Chen, Hairong Yin, Huangying Zhan +33D Scene RepresentationOpen-Vocabulary 3D Segmentation