Traffic rank on Composite
#116 of 180
0.01% of tokens and 0% of requests over 30 days.
AI model
z-ai/glm-5.3-flashx-cheap
Z.ai: GLM 5.3 FlashX (Cheap) is a 1,048,576-token model served on Composite at $0.262 per million tokens in and $0.646 per million tokens out. Ranked #116 of 180 models by traffic on Composite over the last 30 days, taking 0.01% of tokens and 0% of requests. Measured on Composite, 100.0% of its last 4 requests on Composite succeeded, the typical first token arrives in 12463 ms, output runs at about 25.5 tokens per second.
Try in Playground| Model ID | z-ai/glm-5.3-flashx-cheap |
|---|---|
| Provider | cv11 |
| Context window | 1,048,576 tokens |
| Accepts | text, image, video |
| Returns | text |
| Input price | $0.262 per million tokens |
| Output price | $0.646 per million tokens |
| Prompt caching | Not offered for this model |
| Cost on a monthly plan | 5 requests of the daily allowance per call |
| Plans that may spend on it | Basic, Plus, Pro, Ultra, Enterprise |
| Routing endpoints | 2 (base plus 1 routing variant) |
#116 of 180
0.01% of tokens and 0% of requests over 30 days.
100.0%
Across the last 4 requests Composite served for this model.
12463 ms
Median across the same sample.
25.5 tok/s
Median generation rate after the first token.
The same model behind 1 alternative endpoint. Call the id directly to pin a route; the base id above picks for you.
| Route | Model ID | Provider | Price |
|---|---|---|---|
| cheapThink | z-ai/glm-5.3-flashx-cheap:cheapThink |
cv11 | $0.262 per million tokens input · $0.646 per million tokens output |
$0.262 per million tokens input · $0.646 per million tokens output
Input and output are charged at the provider’s list price.
On a monthly plan, one call costs 5 requests of that day's allowance. That needs the Basic or Plus or Pro or Ultra or Enterprise plan; credits pay for it on any plan.
cv11
GLM models balance strong instruction-following with competitive pricing, a reasonable middle ground for most roleplay setups.
With a 1,048,576-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.
It is priced low enough for high-volume daily roleplay without the per-message cost adding up quickly.
Z.ai: GLM 5.3 (Cheap) · Z.ai: GLM 5.2 (Cheap) · Z.ai: GLM 5.1 (Cheap)