GLM 5.1
The cheapest provider for GLM 5.1 is StreamLake at $1.48 per 1M tokens (blended), charging $0.966 per 1M input and $3.036 per 1M output tokens. 13 providers serve it; the most expensive charges 1.45× as much. Context window: 200K tokens. Prices tracked since Oct 3, 2026.
Best price
$1.48
at StreamLake
Input / output
$0.966 / $3.036
per 1M tokens
Providers
13
1.45× spread
7 days
—
best price
30 days
—
best price
Context
200K
tokens
Providers
Active offers ranked by blended price. Bars are relative to the most expensive.
| Provider | Input | Output | Blended | vs cheapest | Context | Uptime 1d | Price since |
|---|---|---|---|---|---|---|---|
StreamLakeCheapest fp8·−31% discount | $0.966 | $3.036 | $1.48 | best | |||
| fp8 | $0.98 | $3.08 | $1.505 | +1% | |||
| fp8 | $1.19 | $3.74 | $1.83 | +23% | |||
| fp8 | $1.26 | $3.96 | $1.935 | +30% | |||
| $1.21 | $4.20 | $1.96 | +32% | ||||
| fp8 | $1.33 | $4.18 | $2.04 | +38% | |||
| fp8 | $1.38 | $4.40 | $2.135 | +44% | |||
| fp8 | $1.40 | $4.40 | $2.15 | +45% | |||
| $1.40 | $4.40 | $2.15 | +45% | ||||
| fp8 | $1.40 | $4.40 | $2.15 | +45% | |||
| fp8 | $1.40 | $4.40 | $2.15 | +45% | |||
| fp8 | $1.40 | $4.40 | $2.15 | +45% | |||
| fp8·−9% discount | $1.40 | $4.40 | $2.15 | +45% |
- StreamLakeCheapest$1.48$0.966 in · $3.036 outbest200K ctx99.4% uptimefp8·−31% discount
- $1.505$0.98 in · $3.08 out+1%198K ctx87.3% uptimefp8
- $1.83$1.19 in · $3.74 out+23%200K ctx99.9% uptimefp8
- $1.935$1.26 in · $3.96 out+30%198K ctx98.5% uptimefp8
- $1.96$1.21 in · $4.20 out+32%198K ctx86.6% uptime
- $2.04$1.33 in · $4.18 out+38%203K ctx99.1% uptimefp8
- $2.135$1.38 in · $4.40 out+44%200K ctx98.9% uptimefp8
- $2.15$1.40 in · $4.40 out+45%198K ctx98.2% uptimefp8
- $2.15$1.40 in · $4.40 out+45%198K ctx100% uptime
- $2.15$1.40 in · $4.40 out+45%198K ctx97.9% uptimefp8
- $2.15$1.40 in · $4.40 out+45%198K ctx98.2% uptimefp8
- $2.15$1.40 in · $4.40 out+45%198K ctx99.9% uptimefp8
- $2.15$1.40 in · $4.40 out+45%200K ctx96.4% uptimefp8·−9% discount
Blended = (3 × input + output) / 4.
Price history
Hover to compare providers at any point in time. Click a legend item to hide it.
USD per 1M tokens · each step is a price change
Not enough history yet
Tracking since Oct 3, 2026. The chart fills in as new crawls arrive every 6 hours.
Other pricing
Cache, reasoning and per-call fees. Token rates per 1M; others per unit.
| Offered by | Cache read |
|---|---|
| StreamLake | $0.179 |
| Chutes | $0.098 |
| SiliconFlow, Phala | $0.60 |
| AtlasCloud | $0.234 |
| Alibaba | $0.247 |
| Novita, Baidu +3 | $0.26 |
| Venice | $0.26 |
Change log
Every recorded change for this model since Oct 3, 2026.
No changes yet
Tracking since Oct 3, 2026. Price changes, new providers and delistings will show up here.
About this model
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
- Released
- Apr 7, 2026
- First tracked
- Oct 3, 2026
- Canonical slug
- z-ai/glm-5.1-20260406
- Weights
- zai-org/GLM-5.1