emergent-modularity

Tag

Cards List
#emergent-modularity

Emergent Modularity in Mixture-of-Experts Models (8 minute read)

TLDR AI · 2026-05-11 Cached

Ai2 releases EMO, a 14B-parameter mixture-of-experts language model trained to develop emergent modularity. It allows using a small subset of experts for specific tasks while maintaining near full-model performance.

0 favorites 0 likes
#emergent-modularity

EMO: Pretraining mixture of experts for emergent modularity

Hugging Face Blog · 2026-05-08 Cached

Allen AI releases EMO, a mixture-of-experts model where modular structure emerges naturally from data, enabling use of just 12.5% of experts for a task while maintaining near full-model performance.

0 favorites 0 likes
#emergent-modularity

EMO: Pretraining Mixture of Experts for Emergent Modularity

Hugging Face Daily Papers · 2026-05-07 Cached

EMO is a Mixture-of-Experts model that enables modular deployment by grouping similar domain tokens with shared experts, achieving performance comparable to standard MoEs while allowing significant expert pruning (25% experts retain 99% performance) without performance degradation.

0 favorites 0 likes
← Back to home

Submit Feedback