expert-expansion

Tag

Cards List
#expert-expansion

Proposed architecture for inferencing sparse MOE models increasing Active parameters using layered + linear decay. Succinct reasoning without any model training or fine tune. [p]

Reddit r/MachineLearning · yesterday

A llama.cpp implementation that expands MoE expert routing beyond native top-K during inference using adaptive thresholds and layer-specific linear decay, without requiring model retraining or fine-tuning.

0 favorites 0 likes
#expert-expansion

Expert expansion with llama.cpp

Reddit r/LocalLLaMA · yesterday

A developer built a custom branch of llama.cpp that implements expert expansion for Mixture-of-Experts models, tested it on Metal, and is seeking cross-platform feedback.

0 favorites 0 likes
← Back to home

Submit Feedback