Google
99.79%Context1M
Max output64K
p50 latency1,849 ms
Throughput145.5 t/s
Input price$0.75
Output price$3.75
No trainingCaching
google/gemini-3.6-flash
Gemini 3.6 Flash delivers higher quality across coding, agentic workflows, and web development with reduced token consumption and fewer model calls compared to previous model iterations.
USD per one million tokens, using current provider data.
Last-hour p50 in milliseconds; lower is better.
Reported availability over the last 24 hours; not a contractual SLA.
Compare current pricing, performance, and privacy across every available provider.
| Provider | Maximum context | Max output | p50 latency | p50 throughput | 24h uptime | Input / 1M tokens | Output / 1M tokens | Privacy |
|---|---|---|---|---|---|---|---|---|
Google | 1M | 64K | 1,849 ms | 145.5 t/s | 99.787% | $0.75 | $3.75 | |
Vertex | 1M | 64K | 1,591 ms | 159.0 t/s | 99.983% | $0.75 | $3.75 |
Available modalities and capabilities across inference providers.
google/gemini-3.6-flashMore options from the same creator or model category.