Llama 4 Scout API Pricing
9 available providers and 9 live price sources tracked, including 0 official provider.
Data snapshot:
30-second decision
- Lowest recorded input price: $0.0166/1M tokens
- Compare official API and relay quotes with their service terms
- Choose this model for repeatable, latency-sensitive, or high-volume workloads where unit cost matters.
Task fit and model switching guide
Derived from official positioning and current ComputeUnion provider and pricing data.
Discount vs official API
Input-price difference for the same model; negative values are above official price.
Official baseline unavailable; discount is not calculated.
Price composition
Input and output remain paired within each provider quote.
Capability radar
Task-fit levels derived from official positioning, not an independent benchmark.
Channel cost and evidence map
X-axis is observed example monthly cost; Y-axis is price-evidence strength.
External capability comparison
Not enough verified related-model data is available for comparison.
Choose a related model for Llama 4 Scout
Compare price, capability, and provider coverage; select a model name for its detail page.
| Model | Official role | Official input/output | External score | Providers | Lowest example monthly cost |
|---|---|---|---|---|---|
| GPT-4o mini | High-volume efficient | $0.15 / $0.6 | — | 14 | $0.01 |
| Llama 3.1 8B Instruct | Not classified | $0.1 / $0.1 | — | 13 | $0.01 |
| DeepSeek V4 Flash | High-volume efficient | $0.14 / $0.28 | 29 | 17 | $0.02 |
- Prices: current official and provider quotes in ComputeUnion, calculated within each provider.
- Capability: official task positioning plus safely matched Artificial Analysis metrics; neither is presented as a ComputeUnion benchmark.
- Evidence: official, live API, platform-submitted, and unreviewed quotes remain visibly separated.
Best-fit tasks
- Best suited to general production workloads.
- Suitable for extraction and classification.
- Suitable for high-volume processing.
When to switch models
- Switch to Llama 4 Maverick when quality and complex reasoning take priority.
- Compare Llama 3.3 70B when lower cost with strong capability matters.
ComputeUnion market view
- ComputeUnion currently tracks 9 providers, including 0 official and 9 live price sources.
Task guidance is generated by fixed rules from official positioning and ComputeUnion market data; it is not an independent benchmark or guarantee.
Provider summary and buying decision
Evidence tier first, then normalized price within each tier.
| Provider / operator | Price evidence | Public operating history | Input / 1M tokens | Output / 1M tokens | API docs | Refund policy | Payment / invoice |
|---|---|---|---|---|---|---|---|
| RunAPIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.0166 | $0.0621 | View docs ↗ | Not provided | alipay · wechatunknown |
| OpenRouterunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.10 | $0.30 | View docs ↗ | Not provided | visa · mastercard · cryptounknown |
| Novita AIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.18 | $0.59 | View docs ↗ | Not provided | visa · mastercardunknown |
| Cloudflare Workers AIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.27 | $0.85 | View docs ↗ | Not provided | visa · mastercardunknown |
| ProAI APIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.3448 | $1.3793 | View docs ↗ | Not provided | alipay · wechatunknown |
| 302.AIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.50 | $0.50 | View docs ↗ | Not provided | alipay · visa · mastercard · amex · unionpayunknown |
| Fireworks AIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.50 | $0.50 | View docs ↗ | Not provided | visa · mastercardunknown |
| YunWu AIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $0.8534 | $3.4138 | View docs ↗ | Not provided | alipay · wechatunknown |
| Together AIunknown | Live API pricePlatform page ↗Observed: | —Unknown | $3.00 | $7.50 | View docs ↗ | Not provided | visa · mastercardunknown |
Platform-submitted information or domain verification does not mean the price was independently checked and is not a ComputeUnion guarantee of pricing, availability, refunds, compliance, invoices, or service quality.
Llama 4 Scout price history
Observed input-price changes across providers over the last 30/90 days.
Unit: $/1M tokens
Estimate monthly cost
The estimate uses the displayed price without assuming discounts or cache hits.
FAQ
How is Llama 4 Scout API priced?
Input and output tokens are billed separately. A complete official input/output baseline is not currently recorded, so provider quotes are normalized per million tokens.
Which provider is cheapest for Llama 4 Scout?
The lowest recorded input price is $0.0166/1M tokens from RunAPI; verify provider terms and the observation date before production use.
How do official and relay prices differ?
Official and relay channels can differ in price, latency, limits, and service terms; evaluate them separately for production traffic.