Back

AI provider / Xiaomi

Xiaomi

A fast-rising Chinese provider centered on MiMo, with a recent push from text reasoning into vision, audio, speech and agent-ready multimodal releases.

Xiaomi now reads as more than a device company experimenting with models. The MiMo site already shows a layered first-party stack: the original MiMo line, MiMo-VL for vision, MiMo-Audio for speech, and the V2 and V2.5 branches that split into Flash, Pro, Omni, ASR and newer distilled variants.

What makes the provider interesting is the bridge between consumer hardware and model deployment. MiMo is not only adding benchmark-style capability, but moving toward the kinds of language, voice and multimodal building blocks that fit assistants, phones, wearables and broader device-native products.

Model catalog

Every release in this provider’s model line.

Each row is a distinct model release — its date, its type and the traits it is known for. Open a source to dig into any of them.

ModelReleasedTypeKey traitsSource
MiMo-V2.5-DFlashJul 2026Cost optimized
  • Distilled flash
  • Lower latency
  • Latest checkpoint
Hugging Face
MiMo-V2.5-ProApr 2026Flagship
  • Top MiMo tier
  • Agentic performance
  • Long context
Xiaomi MiMo
MiMo-V2.5-ASRApr 2026Audio
  • ASR
  • Speech recognition
  • Current speech tier
Hugging Face
MiMo-V2.5Apr 2026Version update
  • New base
  • Reasoning gains
  • Current line
Xiaomi MiMo
MiMo-V2-OmniMar 2026Multimodal
  • Text-audio-vision
  • Assistant focus
  • Omni branch
Xiaomi MiMo
MiMo-V2-ProMar 2026Flagship
  • 1M context
  • Agent workloads
  • Pro tier
Xiaomi MiMo
MiMo-V2-FlashJan 2026Fast tier
  • Lower latency
  • Serving efficiency
  • Reasoning tuned
Xiaomi MiMo
MiMo-AudioSep 2025Audio
  • Speech understanding
  • Voice interaction
  • Audio branch
Xiaomi MiMo
MiMo-VLJun 2025Multimodal
  • Vision-language
  • Image understanding
  • Visual branch
Xiaomi MiMo
MiMoMay 2025Debut
  • Reasoning focus
  • Research release
  • MiMo line
Xiaomi MiMo
Open the model line page