Skip to content
Dashboard

GLM 5.3

GLM 5.3 delivers comprehensive advancements in complex software engineering and agent capabilities. It uses the same base model as GLM-5.2, with all improvements driven by post-training.

Input and output price
50% off
Prices from: Input $0.70, Output $2.20, Per 1M tokens
24h uptime
Loading AI Gateway uptime
import { streamText } from 'ai'
const result = streamText({
model: 'zai/glm-5.3',
prompt: 'Why is the sky blue?'
})
Read docs

Copy link to headingProviders

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the docs for more info. Using a provider means you agree to their terms, listed under Legal.

Promotional pricing ends on September 8, 2026.
Provider
Context
Max Output
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
ZDR
No Training
Regional Inference
Free Tier
Release Date
1M128K3.3 s93 tps
$1.40/M
$4.40/M
Read$0.26/M
08/18/2026
1M1M0.8 s87 tps
$1.20/M
$4/M
Read$0.12/M
08/18/2026
FriendliAI
10% off
Legal:TermsPrivacy
1M1M0.2 s102 tps
$1.26/M
$3.96/M
Read$0.23/M
08/18/2026
1M1M2.4 s15 tps
$1.40/M+1 more
$4.40/M+1 more
Read$0.14/M
US
08/18/2026
1M1M0.8 s147 tps
$1.40/M
$4.40/M
Read$0.26/M
08/18/2026
1M1M1.5 s72 tps
$1.40/M+1 more
$4.40/M+1 more
Read$0.26/M
US
08/18/2026
1M1M0.5 s134 tps
$1.40/M
$4.40/M
Read$0.26/M
08/18/2026
1M1M0.6 s106 tps
$1.40/M
$4.40/M
Read$0.26/M
08/18/2026
1M1M0.1 s295 tps
$1.40/M
$4.40/M
Read$0.26/M
08/18/2026
Legal:TermsPrivacy
1M1M2.1 s199 tps
$0.70/M
$2.20/M
Read$0.13/M
08/18/2026
1M1M0.8 s165 tps
$1.25/M
$4.40/M
Read$0.26/M
08/18/2026
1M128K1.6 s145 tps
$1.40/M
$4.40/M
Read$0.26/M
08/18/2026
1M1M0.6 s67 tps
$1.40/M
$4.40/M
Read$0.26/M
08/18/2026
1M1M0.7 s347 tps
$1.20/M
$4/M
Read$0.20/M
08/18/2026

Copy link to headingPlayground

Try out GLM 5.3 by Z.AI. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.

zai logo
zai logo

GLM 5.3

Copy link to headingUptime

Direct request success rate on AI Gateway and per-provider. Visit the docs for more info.

Copy link to headingThroughput

P50 throughput on live AI Gateway traffic, in tokens per second (TPS). Visit the docs for more info.

Copy link to headingLatency

P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the docs for more info.

Copy link to headingMore models by Z.AI

Model
Context
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
Providers
ZDR
No Training
Free Tier
Release Date
1M1.4 s208 tps
$0.70/M
$2.20/M
Read$0.13/M
digitalocean logo
09/02/2026
1M0.3 s250 tps
$0.07/M
$0.24/M
Read$0.01/M
+1
baseten logo
deepinfra logo
digitalocean logo
+14
08/26/2026
1M0.5 s207 tps
$2.10/M
$6.60/M
Read$0.21/M
alibaba logo
baseten logo
fireworks logo
06/23/2026
1M0.2 s477 tps
$0.70/M+1 more
$2.20/M+1 more
Read$0.11/M
alibaba logo
baseten logo
crusoe logo
+15
06/16/2026
205K1.5 s35 tps
$1.40/M
$4.40/M
Read$0.26/M
deepinfra logo
novita logo
zai logo
04/07/2026
203K0.5 s134 tps
$1/M
$3.20/M
Read$0.20/M
bedrock logo
novita logo
zai logo
02/12/2026

Your use is subject to Z.AI's Terms & Privacy Policies.