XiaomiPremiumPremiumActive

Xiaomi: MiMo-V2.5

MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

Specification
Model ID
xiaomi/mimo-v2.5
Modality
Text+image+audio+video >text
Context
1.1M tokens
Input
$0.140/1M
Output
$0.280/1M
Updated
Sep 20, 2026
Scores
Capabilities
  • Supports tool calling
  • Supports vision inputs
  • Supports long-context workflows
  • Supports deeper reasoning tasks