Method DriftMixture-of-experts routing

Tracked

null experts within token-choice MoE

Improving MoE Compute Efficiency by Composing Weight and Data Sparsity

Mixture-of-experts routing · first seen Jan 21, 2026

current frontier — recent, not yet superseded in the knowledge base

0 papers critique it · 0 beat it on benchmarks

Newer alternatives

Recent methods in the same sub-problem, not yet superseded in the knowledge base.