inclusionAI: Ling 3.0 Flash
The Ling 3.0 Flash is an AI model designed for text-based input and output. It supports reasoning and can utilize tools as part of its processing pipeline. The model handles context lengths up to 262144 tokens and has a maximum completion length of 32768 tokens. For those looking for a balance between price and performance, the Ling 3.0 Flash may be worth shortlisting. With a blended benchmark score of 39.3 across one independent benchmark, it holds a decent standing among its peers. Its pricing structure, at $0.021 per million input tokens and $0.063 per million output tokens, is relatively competitive in the market.
Benchmark results
Independent, published benchmarks. Blended score 39.3 across 1 benchmark, last refreshed 2026-09-23. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 41.1 |
- Model ID
- inclusionai/ling-3.0-flash
- Vendor
- inclusionai
- Released
- July 2026
- Tokenizer
- Other
- Input Modalities
- text
- Output Modalities
- text
- Max Output
- 32,768 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- text only
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $0.02/M input and $0.06/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.01 |
| A month of a busy support chatbot | 5M in / 2M out | $0.23 |
Price & spec history
Tracked daily by PicksByModel since 2026-08-06.
| Date | Input /M | Output /M | Context |
|---|---|---|---|
| 2026-08-07 | $0.02 | $0.06 | 262,144 |
| 2026-08-06 | $0.07 | $0.22 | 131,072 |
Strong choice for
Category rankings
Where inclusionAI: Ling 3.0 Flash places across the 9 categories it ranks in. How we rank →
| # | Category | Score |
|---|---|---|
| #2 | Cheap Bulk InferenceCost · of 25 ranked | 138 |
| #3 | Self-Hosted / LocalCost · of 25 ranked | 118 |
| #4 | Social Media PostsWriting · of 25 ranked | 120 |
| #4 | Voice Assistant BackendVoice · of 25 ranked | 124 |
| #13 | Bulk Data LabelingData · of 25 ranked | 130 |
| #13 | Real-Time ChatLatency · of 25 ranked | 118 |
| #14 | Customer SupportBusiness · of 25 ranked | 128 |
| #14 | Dataset AnnotationResearch · of 25 ranked | 134 |
| #15 | JSON ExtractionData · of 25 ranked | 132 |
Similar models
inclusionAI: Ling 3.0 Flash VL
inclusionAI: Ling 3.0 Flash Fin
inclusionAI: Ling 3.0 Flash VL (free)
inclusionAI: Ling 3.0 Flash Sante (free)
inclusionAI: Ling 3.0 Flash Fin (free)
Quick answers
- How much does inclusionAI: Ling 3.0 Flash cost?
- $0.02 per million input tokens and $0.06 per million output tokens.
- What is inclusionAI: Ling 3.0 Flash's context window?
- 262,144 tokens, roughly 393 pages of text in a single request.
- Does inclusionAI: Ling 3.0 Flash support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can inclusionAI: Ling 3.0 Flash process images?
- No, it is text-only on the input side.