cs.CVOct 5, 2026

Cross-dataset harmonization for robust endoscopic image analysis

Authors: Romil Imtiaz, Dimitris K. Iakovidis

Organizations: Department of Computer Science and Biomedical Informatics, University of Thessaly, Lamia, Greece

Abstract

A significant problem in endoscopic image analysis is that the machine learning (ML) models used for this purpose usually underperform when applied on images acquired from endoscopes that are different from those used to acquire the images of their training set. The main difference of the images originating from different endoscopes is their color distributions, which depend both on the image sensors and the light sources used. Although previous studies have highlighted this challenge, to the best of our knowledge it has not been previously explicitly tackled. This study focuses on this problem and proposes very simple but impactful method. It implements a reference-based image harmonization that reduces global appearance differences between endoscopic datasets. Specifically, it extracts global color statistics from a chosen reference dataset in the CIE-Lab color space and applies a statistical channel-wise transformation to map each target image toward the appearance of the images of the reference dataset. The method is evaluated in the context of polyp detection in both flexible colonoscopy and capsule endoscopy datasets using a dataset-level cross validation protocol. The results indicate that the proposed harmonization consistently improves cross-dataset performance up to 30.7%, outperforming relevant baseline and state-of-the-art methods. The results indicate that a substantial part of the generalization gap is driven by low-level appearance variation that can be mitigated without retraining.

Explore similar work

CardsList
  1. Attribute Retrieving for Open-Vocabulary Endoscopic Compositional Referring Segmentation

    Jul 9, 2026Shun Liu, Nan Xi, Yang Liu +3ColonoscopyImage-Text Pairs

  2. EndoVLM: An Endoscopy Vision-Language Pre-training Model via Anatomy-Guided Sparsity and Progressive Alignment

    Aug 5, 2026Zhenyu Yi, Jianwei Xu, Yue Hu +6Vision-Language AlignmentColonoscopy

  3. Representation-driven Endoscopic Visual Embedding Alignment for Latent Generation

    Aug 7, 2026Francisco Caetano, Tim J. M. Jaspers, Haiko Middeljans +7Recent Vision Foundation ModelsMedical Image Generation