← Back to glossaryGlossary

Mixture of Experts (MoE)

Reviewed 19 July 2026Canonical definitionPart of: AI Agent Cost & Token Terms →

Mixture of Experts is a model architecture that routes each input to a subset of specialised sub-networks rather than processing through the entire model. MoE enables larger, more capable models with lower inference cost.

§01 / QUESTIONSterm: Mixture of Experts (MoE)
Questions

Common questions.

What is Mixture of Experts (MoE)?

Mixture of Experts is a model architecture that routes each input to a subset of specialised sub-networks rather than processing through the entire model.

Why does Mixture of Experts (MoE) matter for AI agents?

MoE enables larger, more capable models with lower inference cost.

Which terms are related to Mixture of Experts (MoE)?

Closely related concepts include Inference Cost, LLM Gateway, Cost per Task (Agent), Token. Each is defined in the Prefactor glossary.

§02 / RELATEDnext: where this fits
Keep reading

Where this fits.

See how every agent performs, and make it better

Prefactor helps teams observe, evaluate, and improve their AI agents in production, across every framework and provider.