MiMo-V2.6-Pro-RL is a 1T-parameter sparse MoE omnimodal model (text, image, video, audio) with 1M-token context, designed as Xiaomi’s flagship agentic model. It scales a unified reinforcement-learning pipeline across coding, general agents, visual tasks, and cybersecurity using groupwise agentic grading and speculative decoding to drive self-improvement and strong benchmark performance
On-demand deployments allow you to use MiMo-V2.6-Pro-RL on dedicated GPUs with Fireworks' high-performance serving stack with high reliability and no rate limits.