Traffic rank on Composite
#183 of 203
0% of tokens and 0% of requests over 30 days.
AI model
qwen/qwen3.8-2.4t-a95b
Qwen: Qwen3.8 2.4T A95B is a 1,048,576-token model served on Composite at $0.640 per million tokens in and $6.00 per million tokens out. Ranked #183 of 203 models by traffic on Composite over the last 30 days, taking 0% of tokens and 0% of requests.
Try in Playground| Model ID | qwen/qwen3.8-2.4t-a95b |
|---|---|
| Provider | cv11 |
| Context window | 1,048,576 tokens |
| Accepts | text |
| Returns | text |
| Input price | $0.640 per million tokens |
| Output price | $6.00 per million tokens |
| Prompt caching | Supported — cached input is 60% off the list rate on pay-as-you-go credits |
| Thinking | Always on for this model |
| Moderated | No |
| Cost on a monthly plan | 20 requests of the daily allowance per call |
| Plans that may spend on it | Plus, Pro, Ultra, Enterprise |
| Routing endpoints | 10 (base plus 9 routing variants) |
#183 of 203
0% of tokens and 0% of requests over 30 days.
The same model behind 9 alternative endpoints. Call the id directly to pin a route; the base id above picks for you.
| Route | Model ID | Provider | Price |
|---|---|---|---|
| Normal | qwen/qwen3.8-2.4t-a95b-normal |
cv11 | $0.640 per million tokens input · $6.00 per million tokens output |
| Normal + cheapThink | qwen/qwen3.8-2.4t-a95b-normal:cheapThink |
cv11 | $0.640 per million tokens input · $6.00 per million tokens output |
| Cheapest provider | qwen/qwen3.8-2.4t-a95b:cheapest-provider |
Novita | $0.640 per million tokens input · $6.00 per million tokens output |
| Cheapest provider + cheapThink | qwen/qwen3.8-2.4t-a95b:cheapest-provider:cheapThink |
Novita | $0.640 per million tokens input · $6.00 per million tokens output |
| cheapThink | qwen/qwen3.8-2.4t-a95b:cheapThink |
cv11 | $0.640 per million tokens input · $6.00 per million tokens output |
| Fast | qwen/qwen3.8-2.4t-a95b:fast |
Alibaba | $0.640 per million tokens input · $6.00 per million tokens output |
| Fast + cheapThink | qwen/qwen3.8-2.4t-a95b:fast:cheapThink |
Alibaba | $0.640 per million tokens input · $6.00 per million tokens output |
| Quality | qwen/qwen3.8-2.4t-a95b:quality |
SiliconFlow | $0.640 per million tokens input · $6.00 per million tokens output |
| Quality + cheapThink | qwen/qwen3.8-2.4t-a95b:quality:cheapThink |
SiliconFlow | $0.640 per million tokens input · $6.00 per million tokens output |
$0.640 per million tokens input · $6.00 per million tokens output
Input and output are charged at the provider’s list price.
On a monthly plan, one call costs 20 requests of that day's allowance. That needs the Plus or Pro or Ultra or Enterprise plan; credits pay for it on any plan.
cv11
Qwen models pair large context windows with solid multilingual ability, useful for roleplay that switches languages or references long lore documents.
With a 1,048,576-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.
It sits in the mid-range on price, a reasonable pick for regular roleplay use without going for the cheapest option available.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
Supplied by cv11, not written by Composite.
Qwen: Qwen3.8 Max Prime · Qwen: Qwen3.8 Omni Flash · Qwen: Qwen3.8 Max (0902)