**MiMo-V2-Flash** is a Mixture-of-Experts (MoE) language model with **309B total parameters** and **15B active parameters**. Designed for high-speed reasoning and agentic workflows, it utilizes a novel hybrid attention architecture and Multi-Token Prediction (MTP) to achieve state-of-the-art performance while significantly reducing inference costs.
# Paste into Claude Code, Cursor, or Codex
Help me set up the Respan gateway: fetch and follow https:/
One provider serves this model through the Respan gateway. Copy a model ID to call that route directly.
| Provider | Input | Output |
|---|---|---|
Featherless featherless_ai/XiaomiMiMo/MiMo-V2-Flash | $0.11/M | $0.28/M |
USD per 1M tokens on Featherless, after gateway discounts.
Total cost from 0 to 1T tokens.
Use MiMo V2 Flash in the tools you already run. Each guide points the app at the Respan gateway, then you pick this model.
Setup guides for more apps are in the integration docs.
Built to meet the security and privacy standards that enterprise and healthcare teams require.
ISO 27001
The internationally recognized standard for information security management.
SOC 2
Secure, compliant management of your data across all of our systems.
GDPR
Operated under GDPR, the world's strictest standard for data privacy.
HIPAA
HIPAA compliant, with a BAA available for healthcare teams.