inclusionai

Ling-3.0-flash

Ling-3.0-flash is an AI model from inclusionai designed for extensive text-based tasks with a context length of 131,072 tokens and support for reasoning and tools. It can process inputs in text format only but offers robust capabilities through advanced reasoning techniques; however, structured output generation is not supported. This model should be considered by users requiring a balance between performance and cost. With a blended benchmark score of 65.2 across three independent benchmarks, Ling-3.0-flash stands as a reliable choice for those looking to integrate an AI with solid coding abilities while managing costs effectively at $0.075 per million input tokens and $0.22 per million output tokens.

Quality Score
86/100
price + capability + benchmarks
Input Price
$0.07
per 1M tokens
Output Price
$0.22
per 1M tokens
Context Window
131,072
tokens

Benchmark results

Independent, published benchmarks. Blended score 65.2 across 3 benchmarks, last refreshed 2026-08-06. How scoring works →

BenchmarkMeasuresScore
AI Index broad capability composite 61.7
AI Index Coding software engineering tasks 83.6
AI Index Agentic multi-step tool-using tasks 48.8
Model ID
inclusionai/ling-3.0-flash
Vendor
inclusionai
Released
July 2026
Tokenizer
Other
Input Modalities
text
Output Modalities
text
Max Output
16,384 tokens
Tool Calling
✓ supported
Structured Output
not supported
Reasoning Mode
✓ supported
Vision
text only
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.07/M input and $0.22/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.05
A month of a busy support chatbot 5M in / 2M out $0.81

Similar models

Quick answers

How much does Ling-3.0-flash cost?
$0.07 per million input tokens and $0.22 per million output tokens.
What is Ling-3.0-flash's context window?
131,072 tokens, roughly 196 pages of text in a single request.
Does Ling-3.0-flash support tool calling?
It supports tool calling, a reasoning mode.
Can Ling-3.0-flash process images?
No, it is text-only on the input side.