Skip to content
Dashboard

Ling 3.0 Flash Sante

Ling-3.0-Flash-Sante is inclusionAI’s language model specialized for health and medicine, built on a Mixture-of-Experts architecture with 124 billion total parameters and approximately 5.1 billion active per token. With a 256K context window and function calling, it supports medical knowledge reasoning, evidence-based retrieval, and complex medical workflows while retaining general reasoning, coding, and agentic capabilities.

Price
Free
24h uptime
Loading AI Gateway uptime
import { streamText } from 'ai'
const result = streamText({
model: 'inclusionai/ling-3.0-flash-sante',
prompt: 'Why is the sky blue?'
})
Read docs

Copy link to headingProviders

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the docs for more info. Using a provider means you agree to their terms, listed under Legal.

Promotional pricing ends on October 4, 2026.
Provider
Context
Max Output
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
ZDR
No Training
Free Tier
Release Date
256K32K0.6 s268 tps
Free
Free
09/04/2026

Copy link to headingPlayground

Try out Ling 3.0 Flash Sante by Inclusionai. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.

I
I

Ling 3.0 Flash Sante

Copy link to headingUptime

Direct request success rate on AI Gateway and per-provider. Visit the docs for more info.

Copy link to headingThroughput

P50 throughput on live AI Gateway traffic, in tokens per second (TPS). Visit the docs for more info.

Copy link to headingLatency

P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the docs for more info.

Copy link to headingMore models by Inclusionai

Model
Context
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
Providers
ZDR
No Training
Free Tier
Release Date
262K0.4 s379 tps
Free
Free
Free
deepinfra logo
novita logo
08/27/2026
256K0.8 s337 tps
$0.06/M
$0.18/M
Read$0.01/M
novita logo
08/06/2026