← All models Model directory

AI model

Qwen: Qwen3.8 Omni Flash

qwen/qwen3.8-omni-flash

Qwen: Qwen3.8 Omni Flash is a 1,000,000-token model served on Composite at $0.0480 per million tokens in and $0.470 per million tokens out.

Try in Playground

Community rating

No community ratings yet.

Sign in to rate this model. Ratings come from accounts that have run requests through Composite.

Specifications

Specifications and pricing as served by Composite
Model IDqwen/qwen3.8-omni-flash
Providercv11
Context window1,000,000 tokens
Acceptstext, image, audio, video
Returnstext
Input price$0.0480 per million tokens
Output price$0.470 per million tokens
Prompt cachingSupported — cached input is 60% off the list rate on pay-as-you-go credits
ModeratedNo
Cost on a monthly plan2 requests of the daily allowance per call
Plans that may spend on itEvery plan, plus pay-as-you-go credits
Routing endpoints2 (base plus 1 routing variant)

Measured on Composite

This model has not carried enough traffic on Composite yet to publish a success rate, a latency figure or a traffic rank. The numbers appear here once it has.

Routing variants

The same model behind 1 alternative endpoint. Call the id directly to pin a route; the base id above picks for you.

RouteModel IDProviderPrice
cheapThink qwen/qwen3.8-omni-flash:cheapThink cv11 $0.0480 per million tokens input · $0.470 per million tokens output

Pricing

$0.0480 per million tokens input · $0.470 per million tokens output

Input and output are charged at the provider’s list price.

On a monthly plan, one call costs 2 requests of that day's allowance.

Provider

cv11

Roleplay fit

Qwen models pair large context windows with solid multilingual ability, useful for roleplay that switches languages or references long lore documents.

Context behavior

With a 1,000,000-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.

Cost for regular use

It is priced low enough for high-volume daily roleplay without the per-message cost adding up quickly.

From the provider

Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...

Supplied by cv11, not written by Composite.