Text / XiaomiMiMo
MiMo V2.6 Pro RL
MiMo V2.6 Pro RL from XiaomiMiMo. Source-based hardware guidance from its published configuration.
EstimatedRepository opened Sep 21, 2026Source checked 9/23/2026Version: 73875d00
LOCALRENTED GPUOPEN WEIGHTS
At a glance
- Parameters
- 1024.22B
- Architecture
- mimo_v2
- Context length
- 1,048,576
- License
- mit
- Software
- SGLang, vLLM
Will it run on your machine?
Save your machine to see a personalized rating and its reasoning.
Add my hardwareWays to run it
SGLang · Linux
Repository-specific command in publisher documentation; confirm dependencies and hardware in the source.
sglang serve --trust-remote-code --model-path XiaomiMiMo/MiMo-V2.6-Pro-RL --tp 16 --dp 2 --enable-dp-attention --mm-enable-dp-encoder --ep 16 --moe-a2a-backend deepep --moe-dense-tp-size 1 --mem-fraction-static 0.7 --max-running-requests 128 \vLLM · Linux
Repository-specific command in publisher documentation; confirm dependencies and hardware in the source.
vllm serve XiaomiMiMo/MiMo-V2.6-Pro-RL --tensor-parallel-size 8 --trust-remote-code --gpu-memory-utilization 0.95 --max-model-len auto --reasoning-parser mimo --tool-call-parser mimo --enable-auto-tool-choice --generation-config vllmExplore its uses
SAFETENSORSTEXT-GENERATION
Keep exploring
Text
Prism MLBonsai 2 · 27B
A compressed 27B-class reasoning model with publisher-provided GGUF packs and a dedicated llama.cpp fork for CUDA, Metal, and CPU.
MODEL SIZE27.36B
Text
DeepSeekDeepSeek V4.1 Flash
A multimodal reasoning model with a compressed key-value cache, published as open weights.
Text
QwenQwen3.8-Flash-Next
An experimental open-weight multimodal model with sparse attention and 262K native context.
MODEL SIZE125B language · 6B active