Traffic rank on Composite
#120 of 180
0.01% of tokens and 0% of requests over 30 days.
AI model
xiaomi/mimo-v2.6-flash
Xiaomi: MiMo-V2.6-Flash is a 1,048,576-token model served on Composite at $0.0448 per million tokens in and $0.280 per million tokens out. Ranked #120 of 180 models by traffic on Composite over the last 30 days, taking 0.01% of tokens and 0% of requests. Measured on Composite, 100.0% of its last 3 requests on Composite succeeded, the typical first token arrives in 8920 ms, output runs at about 76.1 tokens per second.
Try in Playground| Model ID | xiaomi/mimo-v2.6-flash |
|---|---|
| Provider | cv11 |
| Context window | 1,048,576 tokens |
| Accepts | text, image, video, audio |
| Returns | text |
| Input price | $0.0448 per million tokens |
| Output price | $0.280 per million tokens |
| Prompt caching | Supported — cached input is 60% off the list rate on pay-as-you-go credits |
| Moderated | No |
| Cost on a monthly plan | 1 request of the daily allowance per call |
| Plans that may spend on it | Every plan, plus pay-as-you-go credits |
| Routing endpoints | 2 (base plus 1 routing variant) |
#120 of 180
0.01% of tokens and 0% of requests over 30 days.
100.0%
Across the last 3 requests Composite served for this model.
8920 ms
Median across the same sample.
76.1 tok/s
Median generation rate after the first token.
The same model behind 1 alternative endpoint. Call the id directly to pin a route; the base id above picks for you.
| Route | Model ID | Provider | Price |
|---|---|---|---|
| cheapThink | xiaomi/mimo-v2.6-flash:cheapThink |
cv11 | $0.0448 per million tokens input · $0.280 per million tokens output |
$0.0448 per million tokens input · $0.280 per million tokens output
Input and output are charged at the provider’s list price.
On a monthly plan, one call costs 1 request of that day's allowance.
cv11
MiMo models are a newer entrant focused on efficient inference, which tends to show up as lower latency per reply.
With a 1,048,576-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.
It is priced low enough for high-volume daily roleplay without the per-message cost adding up quickly.
MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...
Supplied by cv11, not written by Composite.
Xiaomi: MiMo-V2.6-Pro-UltraSpeed · Xiaomi: MiMo-V2.6-Pro · Xiaomi: MiMo-V2.5-Pro