inclusionai/ling-3.0-flash-sante-free
Ling 3.0 Flash Sante
Ling-3.0-Flash-Sante is inclusionAI’s language model specialized for health and medicine, built on a Mixture-of-Experts architecture with 124 billion total parameters and approximately 5.1 billion active per token. With a 256K context window and function calling, it supports medical knowledge reasoning, evidence-based retrieval, and complex medical workflows while retaining general reasoning, coding, and agentic capabilities.
Price comparison
USD per one million tokens, using current provider data.
Provider uptime
Reported availability over the last 24 hours; not a contractual SLA.
Inference providers
Compare current pricing, performance, and privacy across every available provider.
| Provider | Maximum context | Max output | p50 latency | p50 throughput | 24h uptime | Input / 1M tokens | Output / 1M tokens | Privacy |
|---|---|---|---|---|---|---|---|---|
Novita | 256K | 32K | — | — | 0.000% | $0 | $0 |
What can this model do?
Available modalities and capabilities across inference providers.
Technical profile
- Model type
- Language model
- Release date
- Sep 4, 2026
- Knowledge cutoff
- —
- Regions
- —
- API specifications
- V2, V3, V4
- Temperature control
- Supported
Model identifier
inclusionai/ling-3.0-flash-sante-freeRelated models
More options from the same creator or model category.
