Skip to content

GLM 5.3 (batch)

Z.aitext → textHugging Face OpenRouter

The cheapest provider for GLM 5.3 (batch) is Novita at $0.645 per 1M tokens (blended), charging $0.42 per 1M input and $1.32 per 1M output tokens. 32 providers serve it; the most expensive charges 6.7× as much. Context window: 1M tokens. Prices tracked since Oct 3, 2026.

Best price
$0.645
at Novita
Input / output
$0.42 / $1.32
per 1M tokens
Providers
32
6.7× spread
7 days
—
best price
30 days
—
best price
Context
1M
tokens

Providers

Active offers ranked by blended price. Bars are relative to the most expensive.

  • NovitaCheapest
    $0.645
    $0.42 in · $1.32 out
    best
    1M ctx100% uptimefp8·−70% discount
  • $0.878
    $0.17 in · $3.00 out
    +36%
    256K ctx99.7% uptime
  • $0.901
    $0.179 in · $3.067 out
    +40%
    1M ctx98.4% uptimefp8
  • $1.00
    $0.20 in · $3.40 out
    +55%
    1M ctx99.7% uptimefp8
  • $1.00
    $0.20 in · $3.40 out
    +55%
    1M ctx99.7% uptimefp8
  • $1.01
    $0.22 in · $3.39 out
    +57%
    1M ctx99.9% uptime
  • $1.01
    $0.22 in · $3.39 out
    +57%
    1M ctx99.9% uptime
  • $1.05
    $0.563 in · $2.50 out
    +62%
    1M ctx99.6% uptimefp4·−38% discount
  • $1.075
    $0.70 in · $2.20 out
    +67%
    1M ctx99.7% uptimefp8·−50% discount
  • $1.09
    $0.12 in · $4.00 out
    +69%
    1M ctx99.7% uptime
  • $1.09
    $0.12 in · $4.00 out
    +69%
    1M ctx99.8% uptime
  • $1.16
    $0.24 in · $3.93 out
    +80%
    980K ctx98.6% uptimefp4
  • $1.29
    $0.84 in · $2.64 out
    +100%
    1M ctx99.8% uptime−40% discount
  • $1.30
    $0.60 in · $3.39 out
    +101%
    1M ctx99.5% uptimefp4
  • $1.40
    $0.91 in · $2.86 out
    +117%
    1M ctx99.8% uptime
  • $1.505
    $0.98 in · $3.08 out
    +133%
    1M ctx99.8% uptimefp8·−30% discount
  • $1.68
    $1.053 in · $3.564 out
    +161%
    1M ctx99.9% uptimefp8·−10% discount
  • $1.83
    $1.19 in · $3.74 out
    +183%
    1M ctx99.9% uptime
  • $1.83
    $1.19 in · $3.74 out
    +183%
    1M ctx99.9% uptimefp4·−15% discount
  • $1.935
    $1.26 in · $3.96 out
    +200%
    1M ctx100% uptime−10% discount
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx89.4% uptimefp8
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.3% uptimefp8
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.8% uptimefp4
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.6% uptime
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.8% uptimefp4
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.9% uptime
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.8% uptimenvfp4
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.7% uptimenvfp4
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.3% uptime
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.8% uptimefp4
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.9% uptimefp8
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx100% uptime
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx98.9% uptime
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx98.3% uptime−20% discount
  • $2.15
    $1.40 in · $4.40 out
    +233%
    1M ctx99.9% uptimefp8
  • $3.225
    $2.10 in · $6.60 out
    +400%
    1M ctx99.4% uptimefp8
  • $4.30
    $2.80 in · $8.80 out
    +567%
    1M ctx100% uptime
Blended = (3 × input + output) / 4.

Price history

Hover to compare providers at any point in time. Click a legend item to hide it.

USD per 1M tokens · each step is a price change
Not enough history yet
Tracking since Oct 3, 2026. The chart fills in as new crawls arrive every 6 hours.

Other pricing

Cache, reasoning and per-call fees. Token rates per 1M; others per unit.

Offered byCache read
Novita$0.078
Reka, DigitalOcean$0.169
Morph$0.166
Sail Research, Sail Research (us)$0.15
Wafer, Wafer (us)$0.176
DeepInfra$0.125
SiliconFlow$0.13
InferenceNet, Relace$0.08
Makora$0.19
Phala$0.156
Inceptron$0.175
GMICloud$0.182
AkashML, Friendli$0.234
Alibaba$0.238
Decart$0.196
AtlasCloud, Baidu +9$0.26
BaseTen, Mistral +1$0.14
BaseTen (fast)$0.21
Alibaba (fast)$0.56

Change log

Every recorded change for this model since Oct 3, 2026.

No changes yet
Tracking since Oct 3, 2026. Price changes, new providers and delistings will show up here.

About this model

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
Released
Aug 18, 2026
First tracked
Oct 3, 2026
Canonical slug
z-ai/glm-5.3-20260816
Weights
zai-org/GLM-5.3