Unknown providerLLMactive

Llama 3.1 70B API Pricing

Compare prices from 1 provider, including 0 official. 1 price source is collected automatically.

Data snapshot:

Official input / 1M
Lowest input / 1M$0.90
Maximum input gapNot a service-quality measure
Comparable providers11 live

Llama 3.1 70B provider prices and multipliers

No complete verified official input/output baseline is currently recorded, so provider quotes are shown without multipliers.

Providers: 1From, within billing group: $0.90Lowest observed: $0.90Lowest input multiplier:
Official input baselineOfficial output baselinePrice multipliers do not measure stability, rate limits, refunds, or service quality.
Llama 3.1 70B provider price and evidence comparison
Provider / operatorInput / 1M tokensOutput / 1M tokensvs officialPrice evidencePublic operating historyAPI docsRefund policyPayment / invoice
Fireworks AIOperator type not confirmed$0.90$0.90
In Out
Live API pricePlatform pageObserved: UnknownView docsNot providedVisa · MastercardInvoice support not confirmed

Platform-submitted information or domain verification does not mean the price was independently checked and is not a ComputeUnion guarantee of pricing, availability, refunds, compliance, invoices, or service quality.

Llama 3.1 70B price history

Observed input-price changes across providers over the last 30/90 days.

Unit: $/1M tokens

External capability references

Only source results safely mapped to the official model name are shown.

No safely matched external result is available; adjacent model scores are not substituted.

FAQ

How is Llama 3.1 70B API priced?

Input and output tokens are billed separately. A complete official input/output baseline is not currently recorded, so provider quotes are normalized per million tokens.

Which provider is cheapest for Llama 3.1 70B?

The lowest recorded input price is $0.90/1M tokens from Fireworks AI; verify provider terms and the observation date before production use.

How do official and relay prices differ?

Official and relay channels can differ in price, latency, limits, and service terms; evaluate them separately for production traffic.