Measured on Composite
This model has not carried enough traffic on Composite yet to publish a success rate, a latency figure or a traffic rank. The numbers appear here once it has.
AI model
qwen/qwen3.8-omni-flash
Qwen: Qwen3.8 Omni Flash is a 1,000,000-token model served on Composite at $0.0480 per million tokens in and $0.470 per million tokens out.
Try in Playground| Model ID | qwen/qwen3.8-omni-flash |
|---|---|
| Provider | cv11 |
| Context window | 1,000,000 tokens |
| Accepts | text, image, audio, video |
| Returns | text |
| Input price | $0.0480 per million tokens |
| Output price | $0.470 per million tokens |
| Prompt caching | Supported — cached input is 60% off the list rate on pay-as-you-go credits |
| Moderated | No |
| Cost on a monthly plan | 2 requests of the daily allowance per call |
| Plans that may spend on it | Every plan, plus pay-as-you-go credits |
| Routing endpoints | 2 (base plus 1 routing variant) |
This model has not carried enough traffic on Composite yet to publish a success rate, a latency figure or a traffic rank. The numbers appear here once it has.
The same model behind 1 alternative endpoint. Call the id directly to pin a route; the base id above picks for you.
| Route | Model ID | Provider | Price |
|---|---|---|---|
| cheapThink | qwen/qwen3.8-omni-flash:cheapThink |
cv11 | $0.0480 per million tokens input · $0.470 per million tokens output |
$0.0480 per million tokens input · $0.470 per million tokens output
Input and output are charged at the provider’s list price.
On a monthly plan, one call costs 2 requests of that day's allowance.
cv11
Qwen models pair large context windows with solid multilingual ability, useful for roleplay that switches languages or references long lore documents.
With a 1,000,000-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.
It is priced low enough for high-volume daily roleplay without the per-message cost adding up quickly.
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...
Supplied by cv11, not written by Composite.
Qwen: Qwen3.8 Max (0902) · Qwen: Qwen3.8 Flash · Qwen: Qwen3.8 27B