Before you start
Hugging Face link
https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Flash-RL
Is the model architecture already supported
No, this is a new model architecture
Is the quantization already supported
Not sure
What happens when you load it
Anything else
From the Xiaomi Model Page
Supported sglang - Docker image: lmsysorg/sglang:latest
Supported vLLM - docker pull vllm/vllm-openai:mimov25-cu129
Before you start
Hugging Face link
https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Flash-RL
Is the model architecture already supported
No, this is a new model architecture
Is the quantization already supported
Not sure
What happens when you load it
Anything else
From the Xiaomi Model Page
Supported sglang - Docker image: lmsysorg/sglang:latest
Supported vLLM - docker pull vllm/vllm-openai:mimov25-cu129