Inference.net: Schematron V2 Small
The Inference.net: Schematron V2 Small is a language model that processes text inputs with a context length of up to 128,000 tokens and supports completion tokens up to 4,096 in number. It operates within the realm of text modalities. Consider shortlisting this model if you're working on large-scale text processing tasks or require a moderate-sized output token limit. The pricing for input tokens is $0.05 per million (MTok) and output tokens is $0.23 per MTok, which might be competitive in certain use cases given its performance. However, the lack of benchmark coverage makes it difficult to assess its capabilities against other models; a more comprehensive evaluation will be needed once independent benchmarks become available.
- Model ID
- inference-net/schematron-v2-small
- Vendor
- inference-net
- Released
- September 2026
- Tokenizer
- Other
- Input Modalities
- text
- Output Modalities
- text
- Max Output
- 4,096 tokens
- Tool Calling
- not supported
- Structured Output
- ✓ supported
- Reasoning Mode
- not supported
- Vision
- text only
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $0.05/M input and $0.23/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.04 |
| A month of a busy support chatbot | 5M in / 2M out | $0.71 |
Similar models
Quick answers
- How much does Inference.net: Schematron V2 Small cost?
- $0.05 per million input tokens and $0.23 per million output tokens.
- What is Inference.net: Schematron V2 Small's context window?
- 128,000 tokens, roughly 192 pages of text in a single request.
- Does Inference.net: Schematron V2 Small support tool calling?
- It supports structured output.
- Can Inference.net: Schematron V2 Small process images?
- No, it is text-only on the input side.