GLM 4.5V
The cheapest provider for GLM 4.5V is Novita at $0.90 per 1M tokens (blended), charging $0.60 per 1M input and $1.80 per 1M output tokens. 2 providers serve it at the same price. Context window: 64K tokens. Prices tracked since Oct 3, 2026.
Best price
$0.90
at Novita
Input / output
$0.60 / $1.80
per 1M tokens
Providers
2
1.00× spread
7 days
—
best price
30 days
—
best price
Context
64K
tokens
Providers
Active offers ranked by blended price. Bars are relative to the most expensive.
| Provider | Input | Output | Blended | vs cheapest | Context | Uptime 1d | Price since |
|---|---|---|---|---|---|---|---|
NovitaCheapest fp8 | $0.60 | $1.80 | $0.90 | best | |||
| fp8 | $0.60 | $1.80 | $0.90 | best |
- NovitaCheapest$0.90$0.60 in · $1.80 outbest64K ctx98.8% uptimefp8
- $0.90$0.60 in · $1.80 outbest64K ctx98.6% uptimefp8
Blended = (3 × input + output) / 4.
Price history
Hover to compare providers at any point in time. Click a legend item to hide it.
USD per 1M tokens · each step is a price change
Not enough history yet
Tracking since Oct 3, 2026. The chart fills in as new crawls arrive every 6 hours.
Other pricing
Cache, reasoning and per-call fees. Token rates per 1M; others per unit.
| Offered by | Cache read |
|---|---|
| Novita, Z.AI | $0.11 |
Change log
Every recorded change for this model since Oct 3, 2026.
No changes yet
Tracking since Oct 3, 2026. Price changes, new providers and delistings will show up here.
About this model
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
- Released
- Aug 11, 2025
- First tracked
- Oct 3, 2026
- Canonical slug
- z-ai/glm-4.5v
- Weights
- zai-org/GLM-4.5V