Shih-Hao Hung
1 indexed paper
Recent (6 mo)
1With code
0Influential cites
0Benchmarked
0Publications per year
126
Top categories
Architecture×1
Frequent co-authors
Research Timeline
2026
Decoding the Skew: Distribution-Aware MoE Inference with Adaptive Kernel Dispatch
This paper introduces a distribution-aware framework for modeling and benchmarking Mixture-of-Experts (MoE) inference, showing that the best fused-MoE kernel changes with routing skew and token count, and presents DA-MoE, a GPU-resident kernel-dispatch runtime that improves geomean fused-MoE latency.
Highlighted terms show continued research focus across papers
Papers
cs.AREmpiricalRecentJul 25, 2026
Decoding the Skew: Distribution-Aware MoE Inference with Adaptive Kernel Dispatch
En-Ming Huang, An-Cheng Chang, Bai-Cheng Jeng, Shih-Hao Hung +1 more
This paper introduces a distribution-aware framework for modeling and benchmarking Mixture-of-Experts (MoE) inference, showing that the best fused-MoE kernel changes with routing skew and token count,…
View →