CAPABILITYMOUHN · ACE — ADAPTIVE CORE EXPERTS

Zero wasted experts, proven at scale.

Mixture-of-experts models have a silent problem: as training goes on, many internal experts stop receiving tasks and become dead weight — wasted capacity that's never recovered.

0%experts made useless during training

Traditional mixture-of-experts training lets 20% to 56% of experts stop receiving tasks and become dead weight — capacity that's paid for and never recovered. ACE measures zero wasted experts, across every scale tested.

Measured results

MetricTraditional modelACE
Experts that become useless during training20%–56%0%
Memory use for the same capacitybaselineup to 2× lower
Final quality (same training budget)baselineequal or better, consistent across every scale tested

Confirmed across multiple model scales and multiple independent runs — the zero-waste pattern repeats consistently.

In practice, every expert the model trains actually gets used — so the same training budget buys more real capability, not silent waste.