Skip to content

Ling 3.0 Flash

inclusionAItext → textHugging Face OpenRouter

The cheapest provider for Ling 3.0 Flash is Novita at $0.0315 per 1M tokens (blended), charging $0.021 per 1M input and $0.063 per 1M output tokens. 2 providers serve it; the most expensive charges 2.9× as much. Context window: 256K tokens. Prices tracked since Oct 3, 2026.

Best price
$0.0315
at Novita
Input / output
$0.021 / $0.063
per 1M tokens
Providers
2
2.9× spread
7 days
—
best price
30 days
—
best price
Context
256K
tokens

Providers

Active offers ranked by blended price. Bars are relative to the most expensive.

  • NovitaCheapest
    $0.0315
    $0.021 in · $0.063 out
    best
    256K ctx100% uptime−65% discount
  • $0.09
    $0.06 in · $0.18 out
    +186%
    128K ctx97.8% uptimebf16
Blended = (3 × input + output) / 4.

Price history

Hover to compare providers at any point in time. Click a legend item to hide it.

USD per 1M tokens · each step is a price change
Not enough history yet
Tracking since Oct 3, 2026. The chart fills in as new crawls arrive every 6 hours.

Other pricing

Cache, reasoning and per-call fees. Token rates per 1M; others per unit.

Offered byCache read
Novita$0.0042
DeepInfra$0.012

Change log

Every recorded change for this model since Oct 3, 2026.

No changes yet
Tracking since Oct 3, 2026. Price changes, new providers and delistings will show up here.

About this model

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Released
Jul 23, 2026
First tracked
Oct 3, 2026
Canonical slug
inclusionai/ling-3.0-flash-20260723
Weights
inclusionAI/Ling-3.0-flash