Tag
This paper introduces Mixture of Channel Experts (MoCE), a structured sparse layer that replaces dense pointwise projections in convolutional networks to reduce computational cost while maintaining or improving performance.