Glider: Global and Local Instruction-Driven Expert Router

Pingzhi Li; Prateek Yadav; Jaehong Yoon; Jie Peng; Yi-Lin Sung; Mohit Bansal; Tianlong Chen

doi:10.18653/v1/2025.emnlp-main.319

Glider: Global and Local Instruction-Driven Expert Router

Pingzhi Li, Prateek Yadav, Jaehong Yoon, Jie Peng, Yi-Lin Sung, Mohit Bansal, Tianlong Chen

Abstract

The development of performant pre-trained models has driven the advancement of routing-based expert models tailored to specific tasks. However, these methods often favor generalization over performance on held-in tasks. This limitation adversely impacts practical applicability, as real-world deployments require robust performance across both known and novel tasks. We observe that current token-level routing mechanisms neglect the global semantic context of the input task. To address this, we propose a novel method, Global and Local Instruction Driven Expert Router (GLIDER) that proposes a multi-scale routing mechanism, encompassing a semantic global router and a learned local router. The global router leverages recent LLMs’ semantic reasoning capabilities to generate task-specific instructions from the input query, guiding expert selection across all layers. This global guidance is complemented by a local router that facilitates token-level routing decisions within each module, enabling finer control and enhanced performance on unseen and challenging tasks. Our experiments using T5-based expert models for T0 and FLAN tasks demonstrate that Glider achieves substantially improved held-in performance while maintaining strong generalization on held-out tasks. Additionally, we perform ablations experiments to dive deeper into the components of Glider and plot routing distributions to show that Glider can effectively retrieve the correct expert for held-in tasks while also demonstrating compositional capabilities for held-out tasks. Our experiments highlight the importance of our multi-scale routing that leverages LLM-driven semantic reasoning for MoErging methods.

Anthology ID:: 2025.emnlp-main.319
Volume:: Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing
Month:: November
Year:: 2025
Address:: Suzhou, China
Editors:: Christos Christodoulopoulos, Tanmoy Chakraborty, Carolyn Rose, Violet Peng
Venue:: EMNLP
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 6240–6301
Language:
URL:: https://aclanthology.org/2025.emnlp-main.319/
DOI:: 10.18653/v1/2025.emnlp-main.319
Bibkey:
Cite (ACL):: Pingzhi Li, Prateek Yadav, Jaehong Yoon, Jie Peng, Yi-Lin Sung, Mohit Bansal, and Tianlong Chen. 2025. Glider: Global and Local Instruction-Driven Expert Router. In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, pages 6240–6301, Suzhou, China. Association for Computational Linguistics.
Cite (Informal):: Glider: Global and Local Instruction-Driven Expert Router (Li et al., EMNLP 2025)
Copy Citation:
PDF:: https://aclanthology.org/2025.emnlp-main.319.pdf
Checklist:: 2025.emnlp-main.319.checklist.pdf

PDF Cite Search Checklist Fix data