Z.ai: GLM 5.3 Prime
z-ai/glm-5.3-prime
PaidTool callingReasoning
Price Published
- Input
- $2.80 per million tokens
- Output
- $8.80 per million tokens
Optional features and fallback routes can cost extra. Confirm current terms with the provider.
Specifications Published
- Context window
- 1,000,000 tokens
- Max output
- 131,072 tokens
- Tool calling
- Supported
- Vision input
- Not listed
- Reasoning
- Supported
- Knowledge cutoff
- Not reported
- Listed since
- Sep 23, 2026
Measured by the Club Signed cheapoS reports
Not yet observed No signed request reports for this exact route yet. Missing reports aren't a failure or a zero.
How we measure · Reports come from cheapoS members who share model details.
Try it
- Get an OpenRouter API key. Create one here. Other free tiers are compared on Providers.
- Add it to cheapoS. Open Models → Open OmniRoute, add OpenRouter with your key, then refresh models. Get cheapoS
- Pick this model. Use the ID
z-ai/glm-5.3-primeas a worker or reviewer, and start with a small task you can check.
About this model OpenRouter description
GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...
Similar-priced alternatives with tool calling
- Mistral: Mistral Nemo
mistralai/mistral-nemo$0.02 / $0.03 per M · 131,072 context - inclusionAI: Ling 3.0 Flash VL
inclusionai/ling-3.0-flash-vl$0.02 / $0.06 per M · 262,144 context - inclusionAI: Ling 3.0 Flash
inclusionai/ling-3.0-flash$0.02 / $0.06 per M · 262,144 context - OpenAI: gpt-oss-20b
openai/gpt-oss-20b$0.02 / $0.09 per M · 131,072 context - Meta: Llama 3.1 8B Instruct
meta-llama/llama-3.1-8b-instruct$0.05 / $0.08 per M · 131,072 context - OpenAI: gpt-oss-20b (batch)
openai/gpt-oss-20b:batch$0.02 / $0.11 per M · 131,072 context
Catalog data from the OpenRouter model catalog, retrieved Sep 29, 2026, 12:37 PM UTC. A listing isn't a live availability check.