Traffic rank on Composite
#70 of 180
0.09% of tokens and 0.03% of requests over 30 days.
AI model
xiaomi/mimo-v2.6-pro-ultraspeed
Xiaomi: MiMo-V2.6-Pro-UltraSpeed is a 1,048,576-token model served on Composite at $1.39 per million tokens in and $8.70 per million tokens out. Ranked #70 of 180 models by traffic on Composite over the last 30 days, taking 0.09% of tokens and 0.03% of requests. Measured on Composite, 100.0% of its last 53 requests on Composite succeeded, the typical first token arrives in 29290 ms, output runs at about 1598.3 tokens per second.
Try in Playground| Model ID | xiaomi/mimo-v2.6-pro-ultraspeed |
|---|---|
| Provider | cv11 |
| Context window | 1,048,576 tokens |
| Accepts | text, image, video, audio |
| Returns | text |
| Input price | $1.39 per million tokens |
| Output price | $8.70 per million tokens |
| Prompt caching | Supported — cached input is 60% off the list rate on pay-as-you-go credits |
| Moderated | No |
| Cost on a monthly plan | 36 requests of the daily allowance per call |
| Plans that may spend on it | Plus, Pro, Ultra, Enterprise |
| Routing endpoints | 2 (base plus 1 routing variant) |
#70 of 180
0.09% of tokens and 0.03% of requests over 30 days.
100.0%
Across the last 53 requests Composite served for this model.
29290 ms
Median across the same sample.
1598.3 tok/s
Median generation rate after the first token.
46.7%
Share of prompt tokens served from the provider's cache over 7 days.
The same model behind 1 alternative endpoint. Call the id directly to pin a route; the base id above picks for you.
| Route | Model ID | Provider | Price |
|---|---|---|---|
| cheapThink | xiaomi/mimo-v2.6-pro-ultraspeed:cheapThink |
cv11 | $1.39 per million tokens input · $8.70 per million tokens output |
$1.39 per million tokens input · $8.70 per million tokens output
Input and output are charged at the provider’s list price.
On a monthly plan, one call costs 36 requests of that day's allowance. That needs the Plus or Pro or Ultra or Enterprise plan; credits pay for it on any plan.
cv11
MiMo models are a newer entrant focused on efficient inference, which tends to show up as lower latency per reply.
With a 1,048,576-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.
It sits in the mid-range on price, a reasonable pick for regular roleplay use without going for the cheapest option available.
MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro. Built from the same 1T MiMo-V2.6-Pro checkpoint, it matches the original model in quality while delivering roughly 10x...
Supplied by cv11, not written by Composite.
Xiaomi: MiMo-V2.6-Flash · Xiaomi: MiMo-V2.6-Pro · Xiaomi: MiMo-V2.5-Pro