← All models Model directory

AI model

Xiaomi: MiMo-V2.6-Pro-UltraSpeed

xiaomi/mimo-v2.6-pro-ultraspeed

Xiaomi: MiMo-V2.6-Pro-UltraSpeed is a 1,048,576-token model served on Composite at $1.39 per million tokens in and $8.70 per million tokens out. Ranked #70 of 180 models by traffic on Composite over the last 30 days, taking 0.09% of tokens and 0.03% of requests. Measured on Composite, 100.0% of its last 53 requests on Composite succeeded, the typical first token arrives in 29290 ms, output runs at about 1598.3 tokens per second.

Try in Playground

Community rating

No community ratings yet.

Sign in to rate this model. Ratings come from accounts that have run requests through Composite.

Specifications

Specifications and pricing as served by Composite
Model IDxiaomi/mimo-v2.6-pro-ultraspeed
Providercv11
Context window1,048,576 tokens
Acceptstext, image, video, audio
Returnstext
Input price$1.39 per million tokens
Output price$8.70 per million tokens
Prompt cachingSupported — cached input is 60% off the list rate on pay-as-you-go credits
ModeratedNo
Cost on a monthly plan36 requests of the daily allowance per call
Plans that may spend on itPlus, Pro, Ultra, Enterprise
Routing endpoints2 (base plus 1 routing variant)

Traffic rank on Composite

#70 of 180

0.09% of tokens and 0.03% of requests over 30 days.

Success rate

100.0%

Across the last 53 requests Composite served for this model.

Time to first token

29290 ms

Median across the same sample.

Output speed

1598.3 tok/s

Median generation rate after the first token.

Prompt cache hits

46.7%

Share of prompt tokens served from the provider's cache over 7 days.

Routing variants

The same model behind 1 alternative endpoint. Call the id directly to pin a route; the base id above picks for you.

RouteModel IDProviderPrice
cheapThink xiaomi/mimo-v2.6-pro-ultraspeed:cheapThink cv11 $1.39 per million tokens input · $8.70 per million tokens output

Pricing

$1.39 per million tokens input · $8.70 per million tokens output

Input and output are charged at the provider’s list price.

On a monthly plan, one call costs 36 requests of that day's allowance. That needs the Plus or Pro or Ultra or Enterprise plan; credits pay for it on any plan.

Provider

cv11

Roleplay fit

MiMo models are a newer entrant focused on efficient inference, which tends to show up as lower latency per reply.

Context behavior

With a 1,048,576-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.

Cost for regular use

It sits in the mid-range on price, a reasonable pick for regular roleplay use without going for the cheapest option available.

From the provider

MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro. Built from the same 1T MiMo-V2.6-Pro checkpoint, it matches the original model in quality while delivering roughly 10x...

Supplied by cv11, not written by Composite.