Gmicloud
86.20%Context1M
Max output1M
p50 latency890 ms
Throughput52.0 t/s
Input price$0.13
Output price$0.4
google/gemma-4-26b-a4b-it
Google Gemma 4 26B A4B is a large language model from Google’s Gemma family. It is designed for high-quality text generation, reasoning, and conversational tasks. The model provides strong performance across a wide range of natural language applications, including chat, coding, and content generation.
USD per one million tokens, using current provider data.
Last-hour p50 in milliseconds; lower is better.
Reported availability over the last 24 hours; not a contractual SLA.
Compare current pricing, performance, and privacy across every available provider.
| Provider | Maximum context | Max output | p50 latency | p50 throughput | 24h uptime | Input / 1M tokens | Output / 1M tokens | Privacy |
|---|---|---|---|---|---|---|---|---|
Gmicloud | 1M | 1M | 890 ms | 52.0 t/s | 86.204% | $0.13 | $0.4 | |
Novita | 262.1K | 131.1K | 735 ms | 86.0 t/s | 99.930% | $0.13 | $0.4 | |
Parasail | 262.1K | 131.1K | 974 ms | 15.5 t/s | 99.954% | $0.13 | $0.4 | |
Vertex | 262.1K | 131.1K | 333 ms | 49.0 t/s | 97.459% | $0.15 | $0.6 |
Available modalities and capabilities across inference providers.
google/gemma-4-26b-a4b-itMore options from the same creator or model category.