inclusionai

inclusionAI: Ling 3.1 Flash

Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.

Quality Score
88/100
price + capability + benchmarks
Input Price
Free
per 1M tokens · checked 2026-10-03
Output Price
Free
per 1M tokens · checked 2026-10-03
Context Window
262,144
tokens
Model ID
inclusionai/ling-3.1-flash
Vendor
inclusionai
Released
October 2026
Tokenizer
Other
Input Modalities
text
Output Modalities
text
Max Output
32,768 tokens
Tool Calling
✓ supported
Structured Output
not supported
Reasoning Mode
✓ supported
Vision
text only
Audio
no
Moderated
no

Similar models

Quick answers

Is inclusionAI: Ling 3.1 Flash free to use?
Yes. It is currently listed at no cost per token via OpenRouter, subject to provider rate limits.
What is inclusionAI: Ling 3.1 Flash's context window?
262,144 tokens, roughly 393 pages of text in a single request.
Does inclusionAI: Ling 3.1 Flash support tool calling?
It supports tool calling, a reasoning mode.
Can inclusionAI: Ling 3.1 Flash process images?
No, it is text-only on the input side.

The Model Movers Report

One email every Friday, built from this site's own rankings: the current top five by benchmark score, every model released in the last seven days, and one note worked out from that week's numbers. You can unsubscribe from any issue with one click.