25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Xiaomi / MiMo / MiMo-V2.6-Flash

MiMo-V2.6-Flash

MiMo Flash model for multimodal coding agents and long-context automation

Line

MiMo

Weights

Open

Released

2026-09-22

Context

1.05M

Input

$0.04 / 1M

Output

$0.28 / 1M

Max output

131K

Coverage

not covered yet

Overview

MiMo-V2.6-Flash is a multimodal model from Xiaomi for coding agents and long-context automation. It is part of the MiMo line, succeeding MiMo-V2.5-Pro. It has a context window of 1,048,576 tokens and a maximum output of 131,072 tokens. It accepts audio, image, text, and video, and returns text. It supports reasoning mode and tool calling. Weights are open. Pricing starts at $0.04 per million input tokens and $0.28 per million output tokens, with cached input at $0.0027 per million tokens.

Specs

Released2026-09-22
LineXiaomi · MiMo
WeightsOpen
Context1.05M tokens
Max output131K tokens
Inputaudio, image, text, video
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoffnot stated
API idmimo-v2.6-flash

Pricing

Input$0.04 / 1M tokens
Cached input$0.0027 / 1M tokens
Output$0.28 / 1M tokens

Against MiMo-V2.5-Pro: input $0.40 to $0.04, output $0.80 to $0.28, cached input $0.003 to $0.0027.

Range across the hosts that serve it: up to $0.1692 per 1M input tokens.

Strengths

  • 1M token context window for long documents
  • Accepts audio, image, text, and video inputs
  • Supports reasoning mode and tool calling
  • Open weights for customization
  • 131k token max output

Best for

  • Reach for it for coding agents that need long context
  • Reach for it for multimodal automation pipelines
  • Reach for it for processing long videos or documents
  • Reach for it for tool-calling workflows

How to access

ProviderModel id
AIHubMixmimo-v2.6-flash
ClinePassmimo-v2.6-flash
CrossModelmimo-v2.6-flash
Deep Inframimo-v2.6-flash
DevPass (LLM Gateway)mimo-v2.6-flash
Kilo Gatewaymimo-v2.6-flash
LLM Gatewaymimo-v2.6-flash
NaNmimo-v2.6-flash
NanoGPTmimo-v2.6-flash
OpenCode Gomimo-v2.6-flash
OpenRoutermimo-v2.6-flash
Requestymimo-v2.6-flash

19 hosts serve this model at the price above · the maker's documentation

MiMo: every version

VersionReleasedContextInput
MiMo-V2.6-FlashCurrent2026-09-221.05M$0.04
MiMo-V2.6-Pro2026-09-221.05M$0.43
MiMo-V2.52026-04-221.05M$0.14
MiMo-V2.5-Pro2026-04-221.05M$0.40
MiMo-V2-Omni2026-03-18262K$0.14
MiMo-V2-Pro2026-03-181.05M$0.435

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window size?
The context window is 1,048,576 tokens, and the maximum output is 131,072 tokens.
Does it support image and video input?
Yes, it accepts audio, image, text, and video inputs, and returns text output.
Are the weights open?
Yes, the weights are open. It also supports reasoning mode and tool calling.
Open weightsReasoning1.05M context