Schematron V2 Turbo
Provider:Inference Net
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
Key specs
- Context:128K
- Max output:8.2K
- Tokenizer:Other
- Released:2026-09-12
Token pricing
- Input price:$0.030 / 1M tokens
- Output price:$0.150 / 1M tokens
- Cache read:$0.030 / 1M tokens
- Blended price:$0.060 / 1M tokens
Modalities
Text
Use Cases
- Deep Reasoning & Analysis:Math proofs, logical reasoning, multi-step planning — ideal for research and financial analysis.(≈ $0.0012/call)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Low input + output price: solid fit for high-frequency calls and batch processing.
- Low blended price: budget-friendly for fleet-scale deployment.
FAQ
- What is the token pricing of Schematron V2 Turbo? Input price $0.030 / 1M tokens, output price $0.150 / 1M tokens, blended about $0.060 / 1M tokens. Refer to Inference Net’s official page for the exact rate.
- What context window does Schematron V2 Turbo support? Context window is 128K, max output about 8.2K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is Schematron V2 Turbo best for? Based on its capability and pricing, Schematron V2 Turbo fits best: Deep Reasoning & Analysis. See the Use Cases section for details and cost estimates.
- What input/output modalities does Schematron V2 Turbo support? Supports Text modality.
- Is Schematron V2 Turbo a free model? Schematron V2 Turbo is billed per token, not a free model.
- Which provider offers Schematron V2 Turbo? Schematron V2 Turbo is offered by Inference Net.
More models from Inference Net
Full info:Schematron V2 Turbo