Ling 3.1 Flash from InclusionAI is now available on AI Gateway. The model is free to use through October 13, 2026.
Ling 3.1 Flash is a hybrid reasoning language model with 560B total parameters and 25B active per token. It has a 262K-token context window on AI Gateway.
The model is designed for coding, multi-step analysis, and agents that use tools, including workflows involving long documents, code, and extended task histories.
To try Ling 3.1 Flash, use inclusionai/ling-3.1-flash or inclusionai/ling-3.1-flash-free as the model name:
import { streamText } from 'ai';
const result = streamText({ model: 'inclusionai/ling-3.1-flash-free', prompt: 'Triage the open issues in this repo and group them by theme.',});The standard model ID is free during the promotion and begins billing when it ends. The -free model ID stops serving instead of billing when the free promotion ends.
To use Ling 3.1 Flash in a coding agent, follow the AI Gateway coding agents guide and select inclusionai/ling-3.1-flash as the model.
Try Ling 3.1 Flash in the AI Gateway model playground, or explore the model catalog.
AI Gateway provides one API for calling models, tracking usage and cost, and configuring routing, retries, and failover. You can use an AI Gateway API key or bring your own provider key.