New Model Available: MiMo-V2.6-Flash Xiaomi released MiMo-V2.6-Flash, a natively omnimodal foundation model the company positions for long-context reasoning and agentic workflows across coding, visual understanding, automation, and tool use. The model is listed at $0.14 and $0.28 per million tokens across providers, with read pricing of 0.0028 per million tokens, a 1.05M context window, 3.3s latency, and 57.7 tokens per second throughput, with benchmarks sourced from Artificial Analysis. MiMo-V2.6-Flash is Xiaomi's natively omnimodal foundation model optimized for the best balance of intelligence, efficiency, and cost. It supports long-context reasoning and agentic workflows across coding, visual understanding, automation, and tool use, making it well suited for high-throughput assistants and production applications. Back to Models https://zenmux.ai/models Providers Route requests across multiple providers. Copy a provider slug to set your preference. $0.14 / M tokens $0.28 / M tokens Read: 0.0028 / M tokens Write: - / M tokens1.05M3.3s57.7tps Uptime 24hours Direct request success rate on AI Gateway and per-provider. Throughput 24hours P50 throughput on live AI Gateway traffic, in tokens per second TPS . Latency 24hours P50 time to first token TTFT on live AI Gateway traffic, in milliseconds. Activity Token volume and request traffic to this model over time. Benchmarks Scores on standardized evaluations. Higher percentages are better — and rank percentile shows Metrics sourced from Artificial Analysis https://artificialanalysis.ai/ Apps Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for. View All https://zenmux.ai/analytics/apps Related Models More models from Xiaomi https://zenmux.ai/xiaomi