microsoft / tutel

Tutel MoE: An Optimized Mixture-of-Experts Implementation
MIT License
723 stars 93 forks source link

a bunch of fixes for #167 and #173 #174

Closed ghostplant closed 2 years ago