[ICML 2026] MoSE: Mixture of Slimmable Experts for Efficient and Adaptive Language Models
efficiency language-models mose mixture-of-experts pre-training adaptive-inference llms mixture-of-slimmable-experts post-pretraining
-
Updated
Jul 9, 2026