← All models Model directory

AI model

Z.ai: GLM 5.3 FlashX (Cheap)

z-ai/glm-5.3-flashx-cheap

Z.ai: GLM 5.3 FlashX (Cheap) is a 1,048,576-token model served on Composite at $0.262 per million tokens in and $0.646 per million tokens out. Ranked #116 of 180 models by traffic on Composite over the last 30 days, taking 0.01% of tokens and 0% of requests. Measured on Composite, 100.0% of its last 4 requests on Composite succeeded, the typical first token arrives in 12463 ms, output runs at about 25.5 tokens per second.

Try in Playground

Community rating

No community ratings yet.

Sign in to rate this model. Ratings come from accounts that have run requests through Composite.

Specifications

Specifications and pricing as served by Composite
Model IDz-ai/glm-5.3-flashx-cheap
Providercv11
Context window1,048,576 tokens
Acceptstext, image, video
Returnstext
Input price$0.262 per million tokens
Output price$0.646 per million tokens
Prompt cachingNot offered for this model
Cost on a monthly plan5 requests of the daily allowance per call
Plans that may spend on itBasic, Plus, Pro, Ultra, Enterprise
Routing endpoints2 (base plus 1 routing variant)

Traffic rank on Composite

#116 of 180

0.01% of tokens and 0% of requests over 30 days.

Success rate

100.0%

Across the last 4 requests Composite served for this model.

Time to first token

12463 ms

Median across the same sample.

Output speed

25.5 tok/s

Median generation rate after the first token.

Routing variants

The same model behind 1 alternative endpoint. Call the id directly to pin a route; the base id above picks for you.

RouteModel IDProviderPrice
cheapThink z-ai/glm-5.3-flashx-cheap:cheapThink cv11 $0.262 per million tokens input · $0.646 per million tokens output

Pricing

$0.262 per million tokens input · $0.646 per million tokens output

Input and output are charged at the provider’s list price.

On a monthly plan, one call costs 5 requests of that day's allowance. That needs the Basic or Plus or Pro or Ultra or Enterprise plan; credits pay for it on any plan.

Provider

cv11

Roleplay fit

GLM models balance strong instruction-following with competitive pricing, a reasonable middle ground for most roleplay setups.

Context behavior

With a 1,048,576-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.

Cost for regular use

It is priced low enough for high-volume daily roleplay without the per-message cost adding up quickly.