arXiv AI By Lei Xin, Bin Gu, Peize Li, Zitong Wang, Jianbo Zhao, Changjiang Jiang, Yanyue Xie, Chao Huang, Xuyang Zhao, Zunhai Su, Fanhu Zeng, Zhenglun Kong

UniMoMo: Expert Merging-Based MoE Acceleration for Large Recommendation Models

Read the original on arXiv AI →

arXiv:2608. 08627v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) layers expand recommendation capacity through conditional computation, yet a trained checkpoint still stores and routes over its full expert bank.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.