Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
Inference.net: Schematron V2 Turbo — Inference.net. Context window 128,000 tokens. API pricing (per 1M tokens): Input $0.03, Output $0.15. License: Proprietary.
| Provider | Inference.net |
|---|---|
| API model ID | inference-net/schematron-v2-turbo |
| Tier | Mid |
| Context window | 128,000 tokens |
| Max output | 8,192 tokens |
| License | Proprietary |
| API pricing (per 1M tokens) · Input | $0.03 |
| API pricing (per 1M tokens) · Output | $0.15 |