본문으로 건너뛰기

Mixture of Experts (MoE)

An architecture that routes each token to a few specialized sub-networks, boosting capacity without proportional cost.

관련 리소스