Skip to content

Gemma 4 31B (free)

GoogleFreeimage + text + video → textHugging Face OpenRouter

The cheapest provider for Gemma 4 31B (free) is DeepInfra at $0.153 per 1M tokens (blended), charging $0.09 per 1M input and $0.34 per 1M output tokens. 13 providers serve it; the most expensive charges 5.3× as much. Context window: 256K tokens. Prices tracked since Oct 3, 2026.

Best price
$0.153
at DeepInfra
Input / output
$0.09 / $0.34
per 1M tokens
Providers
13
5.3× spread
7 days
—
best price
30 days
—
best price
Context
256K
tokens

Providers

Active offers ranked by blended price. Bars are relative to the most expensive.

  • DeepInfraCheapest
    turbo
    $0.153
    $0.09 in · $0.34 out
    best
    256K ctx99.3% uptimefp4
  • $0.158
    $0.10 in · $0.33 out
    +3%
    256K ctx94.6% uptime
  • $0.16
    $0.10 in · $0.34 out
    +5%
    256K ctx99.6% uptimefp4
  • $0.18
    $0.12 in · $0.36 out
    +18%
    250K ctx99.4% uptimefp4
  • $0.183
    $0.12 in · $0.37 out
    +20%
    128K ctx95.4% uptimefp4
  • $0.205
    $0.14 in · $0.40 out
    +34%
    256K ctx99.0% uptimebf16
  • $0.205
    $0.14 in · $0.40 out
    +34%
    256K ctx99.8% uptime
  • $0.205
    $0.14 in · $0.40 out
    +34%
    256K ctx83.4% uptimebf16
  • $0.213
    $0.15 in · $0.40 out
    +39%
    256K ctx98.0% uptimefp8
  • $0.213
    $0.15 in · $0.40 out
    +39%
    256K ctx99.0% uptimefp8
  • $0.573
    $0.38 in · $1.15 out
    +275%
    256K ctx99.3% uptime
  • $0.573
    $0.38 in · $1.15 out
    +275%
    128K ctx97.5% uptime
  • $0.813
    $0.75 in · $1.00 out
    +433%
    256K ctx99.8% uptimefp4
  • $0.813
    $0.75 in · $1.00 out
    +433%
    256K ctx81.1% uptimefp8
Blended = (3 × input + output) / 4.

Price history

Hover to compare providers at any point in time. Click a legend item to hide it.

USD per 1M tokens · each step is a price change
Not enough history yet
Tracking since Oct 3, 2026. The chart fills in as new crawls arrive every 6 hours.

Other pricing

Cache, reasoning and per-call fees. Token rates per 1M; others per unit.

Offered byCache read
DeepInfra (turbo), DekaLLM$0.05
CoreWeave$0.10
Venice$0.09
Chutes$0.012
Crusoe$0.14
Parasail$0.06
Io Net$0.19
ModelRun$0.20
SiliconFlow$0.25

Change log

Every recorded change for this model since Oct 3, 2026.

No changes yet
Tracking since Oct 3, 2026. Price changes, new providers and delistings will show up here.

About this model

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Released
Apr 2, 2026
First tracked
Oct 3, 2026
Canonical slug
google/gemma-4-31b-it-20260402
Weights
google/gemma-4-31B-it