Expert Routing

Latest papers 77

All topics
CardsList
  1. DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices

    May 11, 2026Chenyang Song, Weilin Zhao, Xu Han +3Sparse Mixture-of-ExpertsMixture-of-Experts Inference

  2. Hierarchical Mixture-of-Experts with Two-Stage Optimization

    May 8, 2026Gleb Molodtsov, Alexander Miasnikov, Aleksandr BeznosikovExpert Load BalancingMixture of Experts

  3. StrLoRA: Towards Streaming Continual Visual Instruction Tuning for MLLMs

    May 8, 2026Chang Che, Ziqi Wang, Hui Ma +2VLM AdaptationLow-Rank Adaptation

  4. Expert Routing for Communication-Efficient MoE via Finite Expert Banks

    May 6, 2026Mohammad Reza Deylam Salehi, Ali KhalesiMixture of ExpertsInformation-Theoretic Generalization Bounds

  5. Adaptive Inverted-Index Routing for Granular Mixtures-of-Experts

    May 6, 2026Klaus-Rudolf Kladny, Maximilian Mordig, Bernhard Schölkopf +1Mixture-of-Experts InferenceExpert Routing

  6. SceneSelect: Selective Learning for Trajectory Scene Classification and Expert Scheduling

    Apr 27, 2026Xinrun Wang, Deshun Xia, Yuxi Sun +1Multi-Agent Trajectory PredictionExpert Routing

  7. Teacher-Guided Routing for Sparse Vision Mixture-of-Experts

    Apr 23, 2026Masahiro Kada, Ryota Yoshihashi, Satoshi Ikehata +2Adaptive Model RoutingSparse Mixture-of-Experts

  8. Multi-Domain Learning with Global Expert Mapping

    Apr 20, 2026Pourya Shamsolmoali, Masoumeh Zareapoor, Huiyu Zhou +3Vision Foundation Model AdaptationExpert Routing

  9. Layer-wise MoE Routing Locality under Shared-Prefix Code Generation: Token-Identity Decomposition and Compile-Equivalent Fork Redundancy

    Apr 19, 2026Shun-ichiro Hayashi, Daichi Mukunoki, Tetsuya Hoshino +1Expert RoutingEfficient Language Model Inference

  10. OmniMoE: An Efficient MoE by Orchestrating Atomic Experts at Scale

    Feb 5, 2026Jingze Shi, Zhangyang Peng, Yizhang Zhu +3Mixture-of-Experts InferenceLLM Inference Acceleration

  11. Routing-Aware Safety Alignment for Mixture-of-Experts Models

    Feb 4, 2026Jiacheng Liang, Yuhui Wang, Tanqiu Jiang +1Mixture-of-Experts Language ModelsLLM Alignment

  12. Cache-Aware Joint Router Adaptation for Memory-Efficient MoE Inference

    Date pendingZhenhe Wu, Yaping Jin, Qinghua Xing +6Mixture of ExpertsMemory-Efficient Optimization