alibaba/qwen3.8-27b
Qwen3.8 27B
Built on the architectural foundation of Qwen3.5, Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks. Qwen3.8-27B brings these advances to a compact, deployment-friendly dense model: a native vision-language model that understands images and videos, with flexible thinking control, designed to carry complex, multi-step tasks through to completion with greater reliability.
Price comparison
USD per one million tokens, using current provider data.
Provider latency
Last-hour p50 in milliseconds; lower is better.
Provider uptime
Reported availability over the last 24 hours; not a contractual SLA.
Inference providers
Compare current pricing, performance, and privacy across every available provider.
| Provider | Maximum context | Max output | p50 latency | p50 throughput | 24h uptime | Input / 1M tokens | Output / 1M tokens | Privacy |
|---|---|---|---|---|---|---|---|---|
Alibaba | 1M | 131.1K | 1,881 ms | 59.0 t/s | 100.000% | $0.5 | $3 | |
Cerebras | 131.1K | 41K | 465 ms | 678.5 t/s | 100.000% | $0.99 | $1.49 | |
Deepinfra | 262.1K | 262.1K | 681 ms | 31.5 t/s | 100.000% | $0.4 | $3 | |
Morph | 131.1K | 131.1K | 1,216 ms | 73.0 t/s | 100.000% | $0.289 | $2.4 | |
Novita | 1M | 131.1K | 501 ms | 65.5 t/s | 100.000% | $0.42 | $3 | |
Parasail | 262.1K | 262.1K | 647 ms | 113.5 t/s | 100.000% | $0.24 | $2.2 | |
Runinfra | 262.1K | 262.1K | 1,464 ms | 176.0 t/s | 98.907% | $0.1 | $0.4 |
What can this model do?
Available modalities and capabilities across inference providers.
Technical profile
- Model type
- Language model
- Release date
- Aug 14, 2026
- Knowledge cutoff
- —
- Regions
- —
- API specifications
- V2, V3, V4
- Temperature control
- Supported
Model identifier
alibaba/qwen3.8-27bRelated models
More options from the same creator or model category.
