Skip to content

Llama 4 Maverick

Metatext + image → textHugging Face OpenRouter

The cheapest provider for Llama 4 Maverick is DigitalOcean at $0.304 per 1M tokens (blended), charging $0.188 per 1M input and $0.653 per 1M output tokens. 4 providers serve it; the most expensive charges 1.81× as much. Context window: 1M tokens. Prices tracked since Oct 3, 2026.

Best price
$0.304
at DigitalOcean
Input / output
$0.188 / $0.653
per 1M tokens
Providers
4
1.81× spread
7 days
—
best price
30 days
—
best price
Context
1M
tokens

Providers

Active offers ranked by blended price. Bars are relative to the most expensive.

  • DigitalOceanCheapest
    $0.304
    $0.188 in · $0.653 out
    best
    125K ctx99.9% uptime
  • $0.415
    $0.27 in · $0.85 out
    +37%
    1M ctx99.0% uptimefp8
  • $0.513
    $0.35 in · $1.00 out
    +69%
    512K ctx99.7% uptimefp8
  • us-east5
    $0.55
    $0.35 in · $1.15 out
    +81%
    512K ctx
Blended = (3 × input + output) / 4.

Price history

Hover to compare providers at any point in time. Click a legend item to hide it.

USD per 1M tokens · each step is a price change
Not enough history yet
Tracking since Oct 3, 2026. The chart fills in as new crawls arrive every 6 hours.

Other pricing

Cache, reasoning and per-call fees. Token rates per 1M; others per unit.

Offered byCache read
Parasail$0.17

Change log

Every recorded change for this model since Oct 3, 2026.

No changes yet
Tracking since Oct 3, 2026. Price changes, new providers and delistings will show up here.

About this model

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Released
Apr 5, 2025
First tracked
Oct 3, 2026
Canonical slug
meta-llama/llama-4-maverick-17b-128e-instruct
Weights
meta-llama/Llama-4-Maverick-17B-128E-Instruct