Gemini 3.1 Flash-Lite API Pricing
11 available providers and 10 live price sources tracked, including 2 official provider.
Data snapshot:
30-second decision
- Lowest recorded input price: $0.0149/1M tokens
- Compare official API and relay quotes with their service terms
- Choose this model for repeatable, latency-sensitive, or high-volume workloads where unit cost matters.
Task fit and model switching guide
Derived from official positioning and current ComputeUnion provider and pricing data.
Discount vs official API
Input-price difference for the same model; negative values are above official price.
Price composition
Input and output remain paired within each provider quote.
Capability radar
Task-fit levels derived from official positioning, not an independent benchmark.
Channel cost and evidence map
X-axis is observed example monthly cost; Y-axis is price-evidence strength.
External capability comparison for related models
Only evidence-backed related models with the same external metric are compared.
Choose a related model for Gemini 3.1 Flash-Lite
Compare price, capability, and provider coverage; select a model name for its detail page.
| Model | Official role | Official input/output | External score | Providers | Lowest example monthly cost |
|---|---|---|---|---|---|
| GPT-5.4 mini | Not classified | $0.75 / $4.5 | — | 19 | $0.04 |
| Claude Haiku 4.5 | Not classified | $1 / $5 | — | 20 | $0.05 |
| DeepSeek V4 Flash | High-volume efficient | $0.14 / $0.28 | 29 | 17 | $0.02 |
- Prices: current official and provider quotes in ComputeUnion, calculated within each provider.
- Capability: official task positioning plus safely matched Artificial Analysis metrics; neither is presented as a ComputeUnion benchmark.
- Evidence: official, live API, platform-submitted, and unreviewed quotes remain visibly separated.
Best-fit tasks
- Best suited to high-volume processing.
- Best suited to extraction and classification.
- Best suited to translation.
When to switch models
- Switch to Gemini 3.5 Flash when quality and complex reasoning take priority.
- Compare Gemini 3.1 Pro when stronger capability is needed without ignoring cost.
ComputeUnion market view
- ComputeUnion currently tracks 11 providers, including 2 official and 10 live price sources.
- The lowest recorded input price is 94.0% below the official input price; a lower price does not imply better reliability, limits, or refund terms.
Task guidance is generated by fixed rules from official positioning and ComputeUnion market data; it is not an independent benchmark or guarantee.
Provider summary and buying decision
Evidence tier first, then normalized price within each tier.
| Provider / operator | Price evidence | Public operating history | Input / 1M tokens | Output / 1M tokens | API docs | Refund policy | Payment / invoice |
|---|---|---|---|---|---|---|---|
| Google AIunknown | Official price sourcePlatform page ↗Observed: | —Unknown | $0.25 | $1.50 | View docs ↗ | View policy ↗ | Prepay credits · Postpay billingunknown |
| Google Vertex AIunknown | Official price sourcePlatform page ↗Observed: | —Unknown | $0.25 | $1.50 | View docs ↗ | Not provided | visa · mastercard · gcp_creditsunknown |
| RunAPIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.0149 | $0.0894 | View docs ↗ | Not provided | alipay · wechatunknown |
| ProAI APIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.0276 | $0.1655 | View docs ↗ | Not provided | alipay · wechatunknown |
| LaoZhang APIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.0345 | $0.2069 | View docs ↗ | Not provided | alipay · wechatunknown |
| PoloAPIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.0345 | $0.2069 | View docs ↗ | Not provided | alipay · wechatunknown |
| EasyRouterunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.069 | $0.2069 | View docs ↗ | Not provided | alipay · wechatunknown |
| UiUiAPIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.1034 | $0.6207 | View docs ↗ | Not provided | alipay · wechat · visa · mastercardunknown |
| CrazyRouterunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.1375 | $0.825 | View docs ↗ | Not provided | —unknown |
| OpenRouterunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.25 | $1.50 | View docs ↗ | Not provided | visa · mastercard · cryptounknown |
| YunWu AIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.25 | $1.50 | View docs ↗ | Not provided | alipay · wechatunknown |
Platform-submitted information or domain verification does not mean the price was independently checked and is not a ComputeUnion guarantee of pricing, availability, refunds, compliance, invoices, or service quality.
Gemini 3.1 Flash-Lite price history
Observed input-price changes across providers over the last 30/90 days.
Unit: $/1M tokens
Estimate monthly cost
The estimate uses the displayed price without assuming discounts or cache hits.
FAQ
How is Gemini 3.1 Flash-Lite API priced?
Input and output tokens are billed separately. The current recorded official baseline is $0.2500 input / $1.5000 output per 1M tokens.
Which provider is cheapest for Gemini 3.1 Flash-Lite?
The lowest recorded input price is $0.0149/1M tokens from RunAPI; verify provider terms and the observation date before production use.
How do official and relay prices differ?
Official and relay channels can differ in price, latency, limits, and service terms; evaluate them separately for production traffic.