Inference.net · live price
Schematron V2 Turbo API pricing
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
- Input
- $0.03
- per 1M tokens
- Output
- $0.15
- per 1M tokens
- Cached input
- $0.03
- per 1M tokens
- Context
- 128K
- 8K max output
List price from OpenRouter, which passes Inference.net's price through without markup. Refreshed hourly. Released September 12, 2026. Model id inference-net/schematron-v2-turbo.
What Schematron V2 Turbo costs in practice
1,000 chat messages
$0.11
1,000 input + 500 output tokens each
1,000 RAG questions
$0.32
8,000 input + 500 output tokens each
100 coding-agent steps
$0.12
30,000 input + 2,000 output tokens each, no caching
Schematron V2 Turbo price history
Schematron V2 Turbo price by host
- InferenceNet$0.03 in · $0.15 out · 128K ctx · 99.0% uptime 24h
Where to buy Inference.net credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Schematron V2 Turbo cost calculator
Cheaper alternatives to Schematron V2 Turbo
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Qwen: Qwen3.7 Flash | $0.03 | $0.13 |
| DeepSeek: DeepSeek V4 Flash 0423 | $0.042 | $0.084 |
| OpenAI: gpt-oss-20b | $0.018 | $0.09 |
| Mistral: Mistral Small 3 | $0.05 | $0.08 |
| Meta: Llama 3.1 8B Instruct | $0.05 | $0.08 |
| Mistral: Mistral Nemo | $0.019 | $0.03 |
Other Inference.net models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Schematron V2 Small | $0.05 | $0.23 |