Gemma 4 31B API Pricing
3 available providers and 3 live price sources tracked, including 0 official provider.
Data snapshot:
Task fit and model switching guide
Derived from official positioning and current ComputeUnion provider and pricing data.
Discount vs official API
Input-price difference for the same model; negative values are above official price.
Official baseline unavailable; discount is not calculated.
Price composition
Input and output remain paired within each provider quote.
Capability radar
Task-fit levels derived from official positioning, not an independent benchmark.
Channel unit price and evidence map
X-axis is the provider quote per 1M input tokens; Y-axis is price-evidence strength. No usage volume is assumed.
External capability comparison for related models
Only evidence-backed related models with the same external metric are compared.
- Prices: current official and provider quotes in ComputeUnion, calculated within each provider.
- Capability: official task positioning plus safely matched Artificial Analysis metrics; neither is presented as a ComputeUnion benchmark.
- Evidence: official, live API, platform-submitted, and unreviewed quotes remain visibly separated.
Best-fit tasks
- Best suited to coding.
- Suitable for complex reasoning.
- Best suited to general production workloads.
- Best suited to workloads balancing quality and cost.
When to switch models
- Compare Gemma 4 26B A4B for batch, repeatable, or cost-sensitive workloads.
- Switch to Gemini 2.5 Pro when quality and complex reasoning take priority.
- Compare Gemini 3.1 Flash-Lite for batch, repeatable, or cost-sensitive workloads.
ComputeUnion market view
- ComputeUnion currently tracks 3 providers, including 0 official and 3 live price sources.
Task guidance is generated by fixed rules from official positioning and ComputeUnion market data; it is not an independent benchmark or guarantee.
Gemma 4 31B provider prices and multipliers
No complete verified official input/output baseline is currently recorded, so provider quotes are shown without multipliers.
| Provider / operator | Input / 1M tokens | Output / 1M tokens | vs official | Price evidence | Public operating history | API docs | Refund policy | Payment / invoice |
|---|---|---|---|---|---|---|---|---|
| Deep Infraunknown | $0.13 | $0.38 | In —Out — | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | visa · mastercardunknown |
| SambaNovaunknown | $0.22 | $0.59 | In —Out — | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | visa · mastercardunknown |
| Together AIunknown | $0.39 | $0.97 | In —Out — | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | visa · mastercardunknown |
Platform-submitted information or domain verification does not mean the price was independently checked and is not a ComputeUnion guarantee of pricing, availability, refunds, compliance, invoices, or service quality.
Gemma 4 31B price history
Observed input-price changes across providers over the last 30/90 days.
Unit: $/1M tokens
FAQ
How is Gemma 4 31B API priced?
Input and output tokens are billed separately. A complete official input/output baseline is not currently recorded, so provider quotes are normalized per million tokens.
Which provider is cheapest for Gemma 4 31B?
The lowest recorded input price is $0.13/1M tokens from Deep Infra; verify provider terms and the observation date before production use.
How do official and relay prices differ?
Official and relay channels can differ in price, latency, limits, and service terms; evaluate them separately for production traffic.