inclusionai

inclusionAI: Ling 3.0 Flash

The Ling 3.0 Flash is an AI model designed for text-based input and output. It supports reasoning and can utilize tools as part of its processing pipeline. The model handles context lengths up to 262144 tokens and has a maximum completion length of 32768 tokens. For those looking for a balance between price and performance, the Ling 3.0 Flash may be worth shortlisting. With a blended benchmark score of 39.3 across one independent benchmark, it holds a decent standing among its peers. Its pricing structure, at $0.021 per million input tokens and $0.063 per million output tokens, is relatively competitive in the market.

Quality Score
100/100
price + capability + benchmarks
Input Price
$0.02
per 1M tokens · checked 2026-09-23
Output Price
$0.06
per 1M tokens · checked 2026-09-23
Context Window
262,144
tokens

Benchmark results

Independent, published benchmarks. Blended score 39.3 across 1 benchmark, last refreshed 2026-09-23. How scoring works →

BenchmarkMeasuresScore
AI Index broad capability composite 41.1
Model ID
inclusionai/ling-3.0-flash
Vendor
inclusionai
Released
July 2026
Tokenizer
Other
Input Modalities
text
Output Modalities
text
Max Output
32,768 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
✓ supported
Vision
text only
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.02/M input and $0.06/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.01
A month of a busy support chatbot 5M in / 2M out $0.23

Price & spec history

Tracked daily by PicksByModel since 2026-08-06.

DateInput /MOutput /MContext
2026-08-07 $0.02 $0.06 262,144
2026-08-06 $0.07 $0.22 131,072

Strong choice for

Category rankings

Where inclusionAI: Ling 3.0 Flash places across the 9 categories it ranks in. How we rank →

#CategoryScore
#2 Cheap Bulk InferenceCost · of 25 ranked 138
#3 Self-Hosted / LocalCost · of 25 ranked 118
#4 Social Media PostsWriting · of 25 ranked 120
#4 Voice Assistant BackendVoice · of 25 ranked 124
#13 Bulk Data LabelingData · of 25 ranked 130
#13 Real-Time ChatLatency · of 25 ranked 118
#14 Customer SupportBusiness · of 25 ranked 128
#14 Dataset AnnotationResearch · of 25 ranked 134
#15 JSON ExtractionData · of 25 ranked 132

Similar models

Quick answers

How much does inclusionAI: Ling 3.0 Flash cost?
$0.02 per million input tokens and $0.06 per million output tokens.
What is inclusionAI: Ling 3.0 Flash's context window?
262,144 tokens, roughly 393 pages of text in a single request.
Does inclusionAI: Ling 3.0 Flash support tool calling?
It supports tool calling, structured output, a reasoning mode.
Can inclusionAI: Ling 3.0 Flash process images?
No, it is text-only on the input side.

The Model Movers Report

One email every Friday, built from this site's own rankings: the current top five by benchmark score, every model released in the last seven days, and one note worked out from that week's numbers. You can unsubscribe from any issue with one click.