Schematron V2 Turbo
Inference Net · inference-net/schematron-v2-turbo
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
Schematron V2 Turbo ranks #57 of 359 models in the default AI Benchmark Hub order (quality first, then cost). At $0.09 blended per million tokens, it is cheaper than 93% of models with published API pricing. Its 128,000-token context window is 2.0× smaller than the catalog median (262,144 tokens).
Context
Max output: 8192
Pricing
Output / 1M: 0.15
Blend / 1M: 0.09
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.09per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.09 |
| 10,000,000 | $0.90 |
| 100,000,000 | $9.00 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Inference Net
- Schematron V2 Small$0.14 / 1M
Closest API price
- Laguna XS 2.1$0.09 / 1M
- Nova Micro 1.0$0.09 / 1M
- Command R7B (12-2024)$0.09 / 1M
- Mercury 2.5$0.10 / 1M
- MythoMax 13B$0.10 / 1M