ModelRadar
/

MiMo-V2.6-Pro-MOPD

Xiaomi MiMo🇨🇳 CNMajor lab· Sep 27, 2026

Textopen weights

Xiaomi MiMo's multimodal language model combines a sparse mixture of 384 experts with only 42B active parameters and a 1M-token context window. The MOPD2 distillation merges domain-specialized teacher models into a single student, improving agentic tool use and mitigating tool-call repetition on long-horizon tasks such as game development and scientific research.

Summary of the model card

License
mit
Modalities
text → text

1 variant

StandardHF

Sources

HuggingFaceBlog
Deutsch

Xiaomis multimodales Sprachmodell kombiniert ein Sparse-Mixture-of-Experts mit 384 Experten (42B aktiv) und einem Kontextfenster von 1 Million Token. Die MOPD2-Distillation vereint domänenspezifische Lehrermodelle und verbessert agentic Tool-Usage sowie die Reduktion von Tool-Call-Repetition auf langen Aufgabenhorizonten wie Spieleentwicklung und wissenschaftliche Forschung.

Provenance

Detected
Sep 27, 2026
First source
huggingface
Description
Summary of the model card
Editorially reviewed
—