7 repos
inclusionAI/MoBE
Mixture-of-Basis-Experts for Compressing MoE-based LLMs
40
7 commits
thu-nics/MoA
[CoLM'25] The official implementation of the paper <MoA: Mixture of Sparse Attention for Automatic…
160
35 commits
UNITES-Lab/MoE-Quantization
Official code for the paper "Examining Post-Training Quantization for Mixture-of-Experts: A…
31
2 commits
UNITES-Lab/MC-SMoE
[ICLR‘24 Spotlight] Code for the paper "Merge, Then Compress: Demystify Efficient SMoE with Hints…
108
25 commits
lliai/D2MoE
D^2-MoE: Delta Decompression for MoE-based LLMs Compression
86
4 commits
Jab1718/Moe-slices
No description
25
8 commits
JL-Cheng/SERE
[ICLR 2026] SERE: Similarity-Based Expert Re-routing for Efficient Batch Decoding in MoE Models
22
1 commits