Inference.net: Schematron V2 Turbo
by Inference.netInference.net: Schematron V2 Turbo is a large language model from Inference.net. It costs $0.030 per million input tokens and $0.150 per million output tokens. Its context window is 128K tokens.
- Input / 1M tokens
- $0.030
- Output / 1M tokens
- $0.150
- Cached input / 1M
- $0.030
- Context window
- 128K
On repeated prefixes
tokens
Who serves it cheapest
1 hosts serve Schematron V2 Turbo. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| InferenceNetCheapest | $0.030 | $0.150 | 128K | - | 100.0% |
About Schematron V2 Turbo
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
Specifications
| Model ID | inference-net/schematron-v2-turbo |
|---|---|
| Provider | Inference.net |
| Context window | 128K tokens |
| Max output | 8K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | - |
| Open weights | Yes - inference-net/schematron-v2-granite-4.0-h-micro |
| Released | September 12, 2026 |
Cheaper alternatives
Models that cost less than Schematron V2 Turbo while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Inference.net: Schematron V2 Turbo cost?
$0.030 per million input tokens and $0.150 per million output tokens. Cached input reads cost $0.030 per million tokens.
What is the context window of Inference.net: Schematron V2 Turbo?
128K tokens, with up to 8K tokens of output per request.