Mixture-of-Experts Models

Latest papers 297

All topics
CardsList
  1. SAMoRA: Semantic-Aware Mixture of LoRA Experts for Task-Adaptive Learning

    Apr 21, 2026Boyan Shi, Wei Chen, Shuyuan Zhao +4Multi-Task LearningLow-Rank Adaptation

  2. Multi-Domain Learning with Global Expert Mapping

    Apr 20, 2026Pourya Shamsolmoali, Masoumeh Zareapoor, Huiyu Zhou +3Vision Foundation Model AdaptationExpert Routing

  3. Domain-Specialized Object Detection via Model-Level Mixtures of Experts

    Apr 20, 2026Svetlana Pavlitska, Malte Stüven, Beyza Keskin +1Object DetectionMixture-of-Experts Models

  4. Polysemantic Experts, Monosemantic Paths: Routing as Control in MoEs

    Apr 20, 2026Charles Ye, Bo Yuan, Lee SharkeyLLM InterpretabilityNeural Network Interpretability

  5. IMA-MoE: An Interpretable Modality-Aware Mixture-of-Experts Framework for Characterizing the Neurobiological Signatures of Binge Eating Disorder

    Apr 18, 2026Lin Zhao, Qiaohui Gao, Elizabeth Martin +5Multimodal ClassificationMixture-of-Experts Models

  6. Application of a Mixture of Experts-based Foundation Model to the GlueX DIRC Detector

    Apr 17, 2026Cristiano Fanelli, James Giroux, Cole Granger +1High-Energy PhysicsMixture-of-Experts Models

  7. FL-MHSM: Spatially-adaptive Fusion and Ensemble Learning for Flood-Landslide Multi-Hazard Susceptibility Mapping at Regional Scale

    Apr 17, 2026Aswathi Mundayatt, Jaya Sreevalsan-NairMixture-of-Experts Models

  8. Geometric Metrics for MoE Specialization: From Fisher Information to Early Failure Detection

    Apr 16, 2026Dongxin Guo, Jikun Wu, Siu Ming YiuInformation GeometryMixture of Experts

  9. StanceMoE: Mixture-of-Experts Architecture for Stance Detection

    Apr 1, 2026Abdullah Al Shafi, Md. Milon Islam, Sk. Imran Hossain +1Mixture-of-Experts Models

  10. Self-Routing: Parameter-Free Expert Routing from Hidden States

    Apr 1, 2026Jama Hussein Mohamud, Drew Wagner, Mirco RavanelliParameter-Free Mixture-of-Experts RoutingLLM Routing

  11. ExFusion: Efficient Transformer Training via Multi-Experts Fusion

    Mar 30, 2026Jiacheng Ruan, Daize Dong, Xiaoye Qu +5TransformerMixture of Experts

  12. MoE-ACT: Scaling Multi-Task Bimanual Manipulation with Sparse Task-Conditioned Mixture-of-Experts Transformers

    Mar 16, 2026Kangjun Guo, Haichao Liu, Yanji Sun +3Bimanual Robotic ManipulationVisuomotor Policy Learning

  13. Feature-level Interaction Explanations in Multimodal Transformers

    Mar 4, 2026Yeji Kim, Housam Khalifa Bashier Babiker, Mi-Young Kim +1Transformer InterpretabilityFeature Attribution

  14. Modality-Guided Mixture of Structured Experts with Entropy-Triggered Routing for Multimodal Recommendation

    Feb 24, 2026Ji Dai, Quan Fang, DeSheng CaiSparse Mixture-of-ExpertsRecommender Systems

  15. OmniMoE: An Efficient MoE by Orchestrating Atomic Experts at Scale

    Feb 5, 2026Jingze Shi, Zhangyang Peng, Yizhang Zhu +3Mixture-of-Experts InferenceLLM Inference Acceleration

  16. HiMoE-VLA: Hierarchical Mixture-of-Experts for Generalist Vision-Language-Action Policies

    Dec 5, 2025Zhiying Du, Bei Liu, Yaobo Liang +7Visuomotor Policy LearningVision-Language-Action Models

  17. PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference

    Nov 6, 2025Yushu Zhao, Zheng Wang, Minjia ZhangLLM CompressionMixture-of-Experts Models

  18. Variational Mixture of Graph Neural Experts for Alzheimer's Disease Recognition across Frequency Bands in EEG Brain Networks

    Oct 13, 2025Jun-En Ding, Anna Zilverstand, Shihao Yang +2Alzheimer's DiseaseElectroencephalography

  19. SHMoAReg: Spark Deformable Image Registration via Spatial Heterogeneous Mixture of Experts and Attention Heads

    Sep 24, 2025Yuxi Zheng, Jianhui Feng, Tianran Li +2Medical Image RegistrationMixture-of-Experts Models

  20. Macro Graph of Experts for Billion-Scale Multi-Task Recommendation

    Jun 12, 2025Hongyu Yao, Zijin Hong, Hao Chen +6Graph Representation LearningMulti-Task Learning

  21. Multi-Modal Time Series Prediction via Mixture of Modulated Experts

    Date pendingLige Zhang, Ali Maatouk, Jialin Chen +3Multivariate Time Series ForecastingTime Series Forecasting

  22. CMoE: Contrastive Mixture of Experts for Motion Control and Terrain Adaptation of Humanoid Robots

    Date pendingShihao Ma, Hongjin Chen, Zijun Xu +6Humanoid Robot LocomotionTerrain-Aware Robot Locomotion

  23. SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers

    Date pendingShaowen Wang, Ge Zhang, Kairong Luo +6Language Model Scaling LawsTransformer Attention