cs.CVSep 28, 2026

Role-Guided MOE for Encoder-Level Pathology Representation Learning in WSI Classification

Authors: Xinyu Ma, Xing Yang, Hongtao Jin, Guoquan Zhang, Shijie Zhang, Yu Zhang, Xitong Li

Organizations: Artificial Intelligence Research Institute, Shenzhen MSU-BIT University, Shenzhen 518172, China · Shenzhen People’s Hospital, Shenzhen 518020, China

Abstract

Whole slide image classification is a fundamental task in computational pathology, where patch representation quality directly affects downstream aggregation and slide-level discriminability. Pathology foundation models are widely adopted as frozen feature extractors for WSI classification; however, their fixed encoders may produce representations insufficiently adapted to target-specific tissue patterns and discriminative cues. Fine-tuning can improve target adaptation, but introduces a trade-off between pathology-specific representation capacity and adaptation efficiency, particularly in data-scarce settings. To address this, we propose a pathology role-guided mixture-of-experts feed-forward network (MoE-FFN) framework for efficient encoder-level representation learning. We design a two-stage training paradigm to establish and adapt pathology-aware expert specialization. In source-domain expert initialization, pathology-specific priors are distilled from a frozen Virchow2 teacher into a lightweight DINOv2-small student, while role prototypes serve as weak pathological anchors to encourage distinct expert functions. MoE-FFN blocks are introduced into selected high-level transformer layers to provide transformation diversity for heterogeneous pathological patterns. In target-domain adaptation, the initialized experts are refined through asymmetric prototype-guided optimization, enhancing task-relevant positive evidence and separating confusable hard negatives. The resulting encoder extracts offline patch representations that can be directly integrated with standard MIL aggregators. Experiments on the public BRACS dataset and a private PAROTID WSI dataset across five representative backbones demonstrate consistent improvements over the strongest baseline.

Figures & tables

Explore similar work

CardsList
  1. MOOZY: A Patient-First Foundation Model for Computational Pathology

    Mar 27, 2026Yousef Kotp, Vincent Quoc-Huy Trinh, Christopher Pal +1Pathology Foundation ModelsComputational Pathology

  2. Geometry-Aware State Space Model: A New Paradigm for Whole-Slide Image Representation

    May 6, 2026Enhui Chai, Sicheng Chen, Tianyi Zhang +4Whole-Slide ImagesDigital Pathology

  3. LaGuadia: Language-Guided Adaptive Distillation from Pathology Foundation Models

    Jul 13, 2026Gangsu Kim, Won-Ki JeongPathology Foundation ModelsMedical Vision-Language Models