Sparse Mixture-of-Experts

Latest papers 80

All topics
CardsList
  1. Dual-Adaptive SAM3: Hierarchical Routing over Low-Rank Expert Layers for Parameter-Efficient Medical Image Segmentation

    Jun 30, 2026Ying Chen, Jinyue Li, Kun Wang +2VLM AdaptationText-Guided Image Segmentation

  2. TF-MoE: Time-Frequency Mixture-of-Experts for Efficient Speech Separation

    Jun 28, 2026Qinzhe Hu, Chenda Li, Wangyou Zhang +3Sparse Mixture-of-ExpertsSpeech Processing

  3. SARA: Unlocking Multilingual Knowledge in Mixture-of-Experts via Semantically Anchored Routing Alignment

    Jun 24, 2026Tianyu Dong, Yangyang Liu, Jiang Zhou +9Multilingual Language ModelsSparse Mixture-of-Experts

  4. Geometric and Stochastic Analysis of Discontinuities in Sparse Mixture-of-Experts

    Jun 17, 2026Tho Tran Huu, Huu-Tuan Nguyen, Thien-Hai Nguyen +4Mixture of ExpertsSparse Mixture-of-Experts

  5. SoftMoE: Soft Differentiable Routing for Mixture-of-Experts in LLMs

    Jun 16, 2026Mikołaj Zasada, Łukasz Struski, Jacek Tabor +1Mixture-of-Experts Language ModelsSparse Mixture-of-Experts

  6. SPRI: SVD-Partitioned Residual Initialization for Data-Constrained MoE Upcycling

    Jun 15, 2026Weiqiao Shan, Ruixiang Mao, Yuang Li +10Sparse Mixture-of-Experts

  7. FAME: Forecastability-Aware Mixture of Experts for Heterogeneous Time Series Forecasting

    Jun 8, 2026Qianyang Li, Xingjun Zhang, Shaoxun Wang +2Sparse Mixture-of-ExpertsTime Series Forecasting

  8. STAR: Rethinking MoE Routing as Structure-Aware Subspace Learning

    Jun 7, 2026Sumin Park, Noseong ParkSparse Mixture-of-ExpertsExpert Routing

  9. Sparsely gated tiny linear experts

    Jun 5, 2026Simon SchugTransformer FFNsMixture-of-Experts Language Models

  10. Sparse Mixture-of-Experts Reward Models Learn Interpretable and Specialized Experts for Personalized Preference Modeling

    Jun 2, 2026Yifan Wang, Jinyi Mu, Mayank Jobanputra +5Pairwise Preference LearningReward Modeling

  11. Expert-Aware Causal Tracing of Factual Recall in Sparse MoE Language Models

    Jun 2, 2026Yuetian Lu, Ali Modarressi, Yihong Liu +1Mixture-of-Experts Language ModelsSparse Mixture-of-Experts

  12. PRISM: Synergizing Vision Foundation Models via Self-organized Expert Specialization

    Jun 2, 2026Ying Tang, Dong Li, Youjia Zhang +3Vision Foundation ModelsSparse Mixture-of-Experts

  13. Hierarchically Decoupled Mixture-of-Experts for Robust Traffic Sign Recognition in Complex Driving Scenarios

    Jun 1, 2026Mingxiao Wang, Xiaozhen Qu, Bolin Gao +2Sparse Mixture-of-ExpertsAutonomous Driving Perception

  14. ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts

    Jun 1, 2026Heng Zhao, Zilei Shao, Guy Van den Broeck +1Sparse Mixture-of-ExpertsExpert Routing

  15. DAG-MoE: From Simple Mixture to Structural Aggregation in Mixture-of-Experts

    May 31, 2026Jiarui Feng, Hanqing Zeng, Karish Grover +11Mixture-of-Experts Language ModelsSparse Mixture-of-Experts

  16. Eigenvectors of Experts are Training-free Non-collapsing Routers

    May 29, 2026Giang Do, Hung Le, Truyen TranParameter-Free Mixture-of-Experts RoutingMixture of Experts

  17. Beyond Routing: Characterising Expert Tuning and Representation in Vision Mixture-of-Experts

    May 20, 2026Gene Tangtartharakul, Katherine R. StorrsVisual Representation LearningMixture of Experts

  18. UB-SMoE: Universally Balanced Sparse Mixture-of-Experts for Resource-adaptive Federated Fine-tuning of Foundation Models

    May 15, 2026Van-Tuan Tran, Hong-Hanh Nguyen-Le, Marco Ruffini +1Expert Load BalancingSparse Mixture-of-Experts

  19. When Does Sparse MoE Help in Vision? The Role of Backbone Compute Leverage in Sparse Routing

    May 15, 2026Libo Sun, Po-wei Harn, Peixiong He +1Sparse Mixture-of-ExpertsExpert Routing

  20. BEAM: Binary Expert Activation Masking for Dynamic Routing in MoE

    May 14, 2026Juntong Wu, Jialiang Cheng, Qishen Yin +5Sparse Mixture-of-ExpertsMixture-of-Experts Inference