← All models Model directory

AI model

Z.ai: GLM 5.3 Prime

z-ai/glm-5.3-prime

Z.ai: GLM 5.3 Prime is a 1,000,000-token model served on Composite at $0.896 per million tokens in and $8.80 per million tokens out.

Try in Playground

Community rating

— No community ratings yet.

Sign in to rate this model. Ratings come from accounts that have run requests through Composite.

Specifications

Specifications and pricing as served by Composite
Model IDz-ai/glm-5.3-prime
Providercv11
Context window1,000,000 tokens
Acceptstext
Returnstext
Input price$0.896 per million tokens
Output price$8.80 per million tokens
Prompt cachingSupported — cached input is 60% off the list rate on pay-as-you-go credits
ThinkingAlways on for this model
ModeratedNo
Cost on a monthly plan29 requests of the daily allowance per call
Plans that may spend on itPlus, Pro, Ultra, Enterprise
Routing endpoints2 (base plus 1 routing variant)

Measured on Composite

This model has not carried enough traffic on Composite yet to publish a success rate, a latency figure or a traffic rank. The numbers appear here once it has.

Routing variants

The same model behind 1 alternative endpoint. Call the id directly to pin a route; the base id above picks for you.

RouteModel IDProviderPrice
cheapThink z-ai/glm-5.3-prime:cheapThink cv11 $0.896 per million tokens input · $8.80 per million tokens output

Pricing

$0.896 per million tokens input · $8.80 per million tokens output

Input and output are charged at the provider’s list price.

On a monthly plan, one call costs 29 requests of that day's allowance. That needs the Plus or Pro or Ultra or Enterprise plan; credits pay for it on any plan.

Provider

cv11

Roleplay fit

GLM models balance strong instruction-following with competitive pricing, a reasonable middle ground for most roleplay setups.

Context behavior

With a 1,000,000-token context window, it can hold a long-running roleplay or a full lorebook in memory without losing earlier plot details, so it suits multi-session campaigns and character cards with heavy backstory.

Cost for regular use

It sits in the mid-range on price, a reasonable pick for regular roleplay use without going for the cheapest option available.

From the provider

GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...

Supplied by cv11, not written by Composite.