CORTEXA
← Browse

XiuTeng Zhou

1 paper indexed

arxivcs.AI2026-07-02

Generic Expert Coverage for Pruning SparseMixture-of-Experts Language Models

Yongqin Zeng, Sicheng Pan, Jiale Wang, Hai-tao Zheng, Hong-Gee Kim, Chunxia Ma, et al.

Sparsely activated Mixture-of-Experts (MoE) language models contain substantial structured redundancy among routed experts, but pruning them without downstream calibration data remains challenging. Existing expert-pruning methods typically rely on a single aggregated importance s…

View free PDFSource page