Ling 3.0 Flash Sante from inclusionAI is now available on AI Gateway, free to use through October 4.

Ling 3.0 Flash Sante is a health and medicine-focused version of Ling 3.0 Flash. It is a Mixture-of-Experts model with 124B total parameters and about 5.1B active per token, a 256K token context window, and function calling.

The model is built for medical reasoning, professional healthcare tasks, deep research, evidence-based retrieval, and multi-step medical workflows. It retains the base model's general reasoning, coding, and agentic capabilities.

How to use the model during the free period

  • The standard model ID, inclusionai/ling-3.0-flash-sante, is free through October 4 and begins billing when the offer ends.

  • The free model ID, inclusionai/ling-3.0-flash-sante-free, stops serving when the offer ends instead of billing.

Free requests still appear in your spend dashboard and carry a trace, they just cost nothing.

To use Ling 3.0 Flash Sante, set model in the AI SDK:

Use inclusionai/ling-3.0-flash-sante-free if you want the model to stop serving when the offer ends rather than start billing.

To use it in a coding agent, see the coding agents guide, then run vercel ai-gateway coding-agents setup to connect Claude Code, Codex, Cursor, and more, then select inclusionai/ling-3.0-flash-sante in the agent.

Try Ling 3.0 Flash Sante in the model playground, or open the free model page.

AI Gateway provides a unified API for calling models, tracking usage and cost, and configuring retries, failover, and performance optimizations for higher-than-provider uptime. It includes built-in custom reporting, Zero Data Retention support, budgets for API keys, routing rules, and more.

AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference, including on Bring Your Own Key (BYOK) requests.

You can view all language models available on AI Gateway.

Read more