Ling-3.0-flash-Sante
AvailableA health- and medicine-domain-tuned variant of Ant Group inclusionAI's Ling-3.0-flash, launched September 4 2026 (model id inclusionai/ling-3.0-flash-sante). Same efficient sparse Mixture-of-Experts base — 124B total parameters, ~5.1B active per token — post-trained for medical knowledge reasoning, clinical safety, evidence-based retrieval, and long-horizon medical tasks, while inclusionAI reports it retains strong general reasoning, coding, and agentic ability. Text-in / text-out only (no vision), with a 262,144-token (256K) context and up to 32,768 output tokens; supports reasoning and tool / function calling. Available first via hosted serverless API — Novita, OpenRouter (inclusionai/ling-3.0-flash-sante), and Vercel AI Gateway — with a time-limited free window at launch (free through Oct 4 on Vercel AI Gateway). Positioned as a developer API for research, retrieval, summarization, and workflow assistance, explicitly not a medical device or a substitute for clinical judgment; no public benchmark table at launch, so treat domain claims as unverified. Like the Fin variant, Sante-specific open weights were not confirmed posted at launch (the base Ling-3.0-flash family ships open-weight, MIT), so treat the open-weight status as announced-but-unconfirmed for this variant.
Specifications
- License
- Open weights · Sante-specific weights not confirmed posted at launch (hosted API-first); base Ling-3.0-flash family is MIT
- Weights
- Not released
- Architecture
- Mixture-of-Experts
- Parameters
- 124B · 5.1B active
- Context window
- 262K tokens
- Max output
- 33K tokens
- Knowledge cutoff
- —
- Price (in / out, $/M)
- —
- Modalities
- TextCode
Benchmarks
No benchmark scores recorded yet. Spotted some? Submit a correction.
Vendor-reported figures are claims until independently verified. See methodology.