Inference.net: Schematron V2 Turbo
The Inference.net Schematron V2 Turbo is a text-based AI model that handles input modalities limited to text and has a maximum context length of 128,000 tokens. It does not support tools, reasoning, or structured output. For those prioritizing price above performance, the Schematron V2 Turbo may be worth considering due to its competitive pricing structure: $0.03 per million input tokens and $0.15 per million output tokens. However, it is currently missing independent benchmark coverage, making a full assessment of its capabilities challenging.
- Model ID
- inference-net/schematron-v2-turbo
- Vendor
- inference-net
- Released
- September 2026
- Tokenizer
- Other
- Input Modalities
- text
- Output Modalities
- text
- Max Output
- 8,192 tokens
- Tool Calling
- not supported
- Structured Output
- ✓ supported
- Reasoning Mode
- not supported
- Vision
- text only
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $0.03/M input and $0.15/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.02 |
| A month of a busy support chatbot | 5M in / 2M out | $0.45 |
Similar models
Quick answers
- How much does Inference.net: Schematron V2 Turbo cost?
- $0.03 per million input tokens and $0.15 per million output tokens.
- What is Inference.net: Schematron V2 Turbo's context window?
- 128,000 tokens, roughly 192 pages of text in a single request.
- Does Inference.net: Schematron V2 Turbo support tool calling?
- It supports structured output.
- Can Inference.net: Schematron V2 Turbo process images?
- No, it is text-only on the input side.