Glossary · Models & inference

MoE (Mixture of Experts)

An architecture with multiple expert subnetworks and a learned router that selects a subset for each input unit, often each token. Sparse activation can increase total parameter capacity without using every expert on every forward pass.

Why it matters

Compute, memory, communication, routing balance, and quality depend on the specific architecture and serving system.

Common confusion

Product names do not prove an MoE architecture unless the model developer discloses it.

Related terms

Browse the learning paths to see this term in context — every lesson is free to read.