← All models Model directory

AI model

Qwen: Qwen3.8 2.4T A95B

qwen/qwen3.8-2.4t-a95b

Qwen: Qwen3.8 2.4T A95B is a 1,048,576-token model served on Composite at $0.640 per million tokens in and $6.00 per million tokens out. Ranked #183 of 203 models by traffic on Composite over the last 30 days, taking 0% of tokens and 0% of requests.

Try in Playground

Community rating

— No community ratings yet.

Sign in to rate this model. Ratings come from accounts that have run requests through Composite.

Specifications

Specifications and pricing as served by Composite
Model IDqwen/qwen3.8-2.4t-a95b
Providercv11
Context window1,048,576 tokens
Acceptstext
Returnstext
Input price$0.640 per million tokens
Output price$6.00 per million tokens
Prompt cachingSupported — cached input is 60% off the list rate on pay-as-you-go credits
ThinkingAlways on for this model
ModeratedNo
Cost on a monthly plan20 requests of the daily allowance per call
Plans that may spend on itPlus, Pro, Ultra, Enterprise
Routing endpoints10 (base plus 9 routing variants)

Traffic rank on Composite

#183 of 203

0% of tokens and 0% of requests over 30 days.

Routing variants

The same model behind 9 alternative endpoints. Call the id directly to pin a route; the base id above picks for you.

RouteModel IDProviderPrice
Normal qwen/qwen3.8-2.4t-a95b-normal cv11 $0.640 per million tokens input · $6.00 per million tokens output
Normal + cheapThink qwen/qwen3.8-2.4t-a95b-normal:cheapThink cv11 $0.640 per million tokens input · $6.00 per million tokens output
Cheapest provider qwen/qwen3.8-2.4t-a95b:cheapest-provider Novita $0.640 per million tokens input · $6.00 per million tokens output
Cheapest provider + cheapThink qwen/qwen3.8-2.4t-a95b:cheapest-provider:cheapThink Novita $0.640 per million tokens input · $6.00 per million tokens output
cheapThink qwen/qwen3.8-2.4t-a95b:cheapThink cv11 $0.640 per million tokens input · $6.00 per million tokens output
Fast qwen/qwen3.8-2.4t-a95b:fast Alibaba $0.640 per million tokens input · $6.00 per million tokens output
Fast + cheapThink qwen/qwen3.8-2.4t-a95b:fast:cheapThink Alibaba $0.640 per million tokens input · $6.00 per million tokens output
Quality qwen/qwen3.8-2.4t-a95b:quality SiliconFlow $0.640 per million tokens input · $6.00 per million tokens output
Quality + cheapThink qwen/qwen3.8-2.4t-a95b:quality:cheapThink SiliconFlow $0.640 per million tokens input · $6.00 per million tokens output

Pricing

$0.640 per million tokens input · $6.00 per million tokens output

Input and output are charged at the provider’s list price.

On a monthly plan, one call costs 20 requests of that day's allowance. That needs the Plus or Pro or Ultra or Enterprise plan; credits pay for it on any plan.

Provider

cv11

Roleplay fit

Qwen models pair large context windows with solid multilingual ability, useful for roleplay that switches languages or references long lore documents.

Context behavior

With a 1,048,576-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.

Cost for regular use

It sits in the mid-range on price, a reasonable pick for regular roleplay use without going for the cheapest option available.

From the provider

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

Supplied by cv11, not written by Composite.