inclusionai

inclusionAI: Ling 3.0 Flash Sante

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...

Quality Score
94/100
price + capability + benchmarks
Input Price
$0.04
per 1M tokens · checked 2026-10-09
Output Price
$0.12
per 1M tokens · checked 2026-10-09
Context Window
262,144
tokens
Model ID
inclusionai/ling-3.0-flash-sante
Vendor
inclusionai
Released
September 2026
Tokenizer
Other
Input Modalities
text
Output Modalities
text
Max Output
32,768 tokens
Tool Calling
✓ supported
Structured Output
not supported
Reasoning Mode
✓ supported
Vision
text only
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.04/M input and $0.12/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.03
A month of a busy support chatbot 5M in / 2M out $0.46

Similar models

Quick answers

How much does inclusionAI: Ling 3.0 Flash Sante cost?
$0.04 per million input tokens and $0.12 per million output tokens.
What is inclusionAI: Ling 3.0 Flash Sante's context window?
262,144 tokens, roughly 393 pages of text in a single request.
Does inclusionAI: Ling 3.0 Flash Sante support tool calling?
It supports tool calling, a reasoning mode.
Can inclusionAI: Ling 3.0 Flash Sante process images?
No, it is text-only on the input side.

The Model Movers Report

One email every Friday, built from this site's own rankings: the current top five by benchmark score, every model released in the last seven days, and one note worked out from that week's numbers. You can unsubscribe from any issue with one click.