Bedrock
100.00%Context128K
Max output8.2K
p50 latency195 ms
Throughput209.0 t/s
Input price$0.72
Output price$0.72
Zero retentionNo training
meta/llama-3.3-70b
Where performance meets efficiency. This model supports high-performance conversational AI designed for content creation, enterprise applications, and research, offering advanced language understanding capabilities, including text summarization, classification, sentiment analysis, and code generation.
USD per one million tokens, using current provider data.
Last-hour p50 in milliseconds; lower is better.
Reported availability over the last 24 hours; not a contractual SLA.
Compare current pricing, performance, and privacy across every available provider.
| Provider | Maximum context | Max output | p50 latency | p50 throughput | 24h uptime | Input / 1M tokens | Output / 1M tokens | Privacy |
|---|---|---|---|---|---|---|---|---|
Bedrock | 128K | 8.2K | 195 ms | 209.0 t/s | 100.000% | $0.72 | $0.72 |
Available modalities and capabilities across inference providers.
meta/llama-3.3-70bMore options from the same creator or model category.