cs.CVAug 21, 2026

VisTa3D: A Dataset and Benchmark for Thin Object Reconstruction from Vision, Tactile, and 3D Point Clouds

Authors: Shania Guo, Yeongsik Seo, Andrew Fu, Mei Hao, Iris Xia, Jiwon Jenny Lee, Xinyi Mary Xie, Hyoungseob Park, +2 more

Abstract

State-of-the-art 3D reconstruction models, whether from visual, range, or both, tend to underperform on thin objects. This is partially due to the small amount of space such objects occupy in RGB images and in 3D point clouds. To test the extent of their errors, we collected the first thin object dataset comprising of synchronized RGB images, depth maps, and tactile response maps, where each frame is associated with inertial measurements, camera pose and calibration, and groundtruth depth and segmentation maps obtained from laser scanning of thin objects. We hypothesize that tactile data can aid in the reconstruction of thin objects as their response maps provide local shape and deformation information. Our dataset, termed VisTa3D, comprises of 387 scenes covering 70 thin objects over 17 environments. We benchmarked current 3D reconstruction models on VisTa3D and found that, indeed, they exhibit low fidelity on thin objects. To test if tactile data can help, we introduce the first visual-range-tactile 3D reconstruction model as a baseline. Code and data: https://huggingface.co/datasets/shaniaguo/VisTa3D.

Explore similar work

CardsList
  1. MessyKitchens: Contact-rich object-level 3D scene reconstruction

    Mar 17, 2026Junaid Ahmed Ansari, Ran Ding, Fabio Pizzati +13D Reconstruction

  2. 3DReflecNet: A Large-Scale Dataset for 3D Reconstruction of Reflective, Transparent, and Low-Texture Objects

    May 11, 2026Zhicheng Liang, Haoyi Yu, Boyan Li +63D ReconstructionMulti-View 3D Reconstruction

  3. TouchTherm: Building Multimodal Digital Twins of Objects for Tactile and Thermal Rendering

    Oct 1, 2026Yitao Zhang, Hong Ying, Haoran Guo +33D Mesh ReconstructionDigital Twins