Skip to content

Schematron V2 Turbo

Inference.nettext → textHugging Face OpenRouter

The cheapest provider for Schematron V2 Turbo is InferenceNet at $0.06 per 1M tokens (blended), charging $0.03 per 1M input and $0.15 per 1M output tokens. It is served by a single provider. Context window: 125K tokens. Prices tracked since Oct 3, 2026.

Best price
$0.06
at InferenceNet
Input / output
$0.03 / $0.15
per 1M tokens
Providers
1
no spread
7 days
—
best price
30 days
—
best price
Context
125K
tokens

Providers

Active offers ranked by blended price. Bars are relative to the most expensive.

  • InferenceNetCheapest
    $0.06
    $0.03 in · $0.15 out
    best
    125K ctx99.8% uptime
Blended = (3 × input + output) / 4.

Price history

Hover to compare providers at any point in time. Click a legend item to hide it.

USD per 1M tokens · each step is a price change
Not enough history yet
Tracking since Oct 3, 2026. The chart fills in as new crawls arrive every 6 hours.

Other pricing

Cache, reasoning and per-call fees. Token rates per 1M; others per unit.

Offered byCache read
InferenceNet$0.03

Change log

Every recorded change for this model since Oct 3, 2026.

No changes yet
Tracking since Oct 3, 2026. Price changes, new providers and delistings will show up here.

About this model

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
Released
Sep 12, 2026
First tracked
Oct 3, 2026
Canonical slug
inference-net/schematron-v2-turbo-20260902
Weights
inference-net/schematron-v2-granite-4.0-h-micro