DeepSeekLLMactive

DeepSeek V4 Flash API Pricing

17 available providers and 17 live price sources tracked, including 1 official provider.

Data snapshot:

30-second decision

  • Lowest recorded input price: $0.0158/1M tokens
  • Compare official API and relay quotes with their service terms
  • Choose this model for repeatable, latency-sensitive, or high-volume workloads where unit cost matters.
Official input / 1M$0.1400
Lowest input / 1M$0.0158
Maximum input gap88.7%Not a service-quality measure
Comparable providers1717 live

Task fit and model switching guide

Derived from official positioning and current ComputeUnion provider and pricing data.

Official positioningOfficially positioned for efficient, high-volume workloads.

Discount vs official API

Input-price difference for the same model; negative values are above official price.

Price composition

Input and output remain paired within each provider quote.

Capability radar

Task-fit levels derived from official positioning, not an independent benchmark.

Channel cost and evidence map

X-axis is observed example monthly cost; Y-axis is price-evidence strength.

External capability comparison for related models

Only evidence-backed related models with the same external metric are compared.

Artificial Analysis

Choose a related model for DeepSeek V4 Flash

Compare price, capability, and provider coverage; select a model name for its detail page.

ModelOfficial roleOfficial input/outputExternal scoreProvidersLowest example monthly cost
Hy3High-volume efficient$0.137931 / $0.551724416$0.24
LongCat 2.0Balanced workhorse$0.75 / $2.95333$0.54
GPT-5.4 miniNot classified$0.75 / $4.519$0.04
Sources and verification method
  • Prices: current official and provider quotes in ComputeUnion, calculated within each provider.
  • Capability: official task positioning plus safely matched Artificial Analysis metrics; neither is presented as a ComputeUnion benchmark.
  • Evidence: official, live API, platform-submitted, and unreviewed quotes remain visibly separated.

Best-fit tasks

  • Best suited to high-volume processing.
  • Suitable for extraction and classification.
  • Best suited to cost-sensitive workloads.

When to switch models

  • Switch to DeepSeek V4 Pro when quality and complex reasoning take priority.

ComputeUnion market view

  • ComputeUnion currently tracks 17 providers, including 1 official and 17 live price sources.
  • The lowest recorded input price is 88.7% below the official input price; a lower price does not imply better reliability, limits, or refund terms.
Example usage1M input + 0.2M output
Official monthly cost$0.196
Lowest recorded monthly cost$0.0221

Task guidance is generated by fixed rules from official positioning and ComputeUnion market data; it is not an independent benchmark or guarantee.

Provider summary and buying decision

Evidence tier first, then normalized price within each tier.

Market low: $0.0158Lowest observed: $0.0158
DeepSeek V4 Flash provider price and evidence comparison
Provider / operatorPrice evidencePublic operating historyInput / 1M tokensOutput / 1M tokensAPI docsRefund policyPayment / invoice
DeepSeekunknownOfficial price sourcePlatform pageObserved: Unknown$0.14$0.28View docsNot providedvisa · mastercard · alipay · wechatunknown
RunAPIunknownLive API pricePlatform pageObserved: Unknown$0.0158$0.0315View docsNot providedalipay · wechatunknown
EasyRouterunknownLive API pricePlatform pageObserved: Unknown$0.0193$0.0386View docsNot providedalipay · wechatunknown
LaoZhang APIunknownLive API pricePlatform pageObserved: Unknown$0.0193$0.0386View docsNot providedalipay · wechatunknown
TreeRouterunknownLive API pricePlatform pageObserved: Unknown$0.0414$0.0828View docsNot providedalipay · wechatunknown
Deep InfraunknownLive API pricePlatform pageObserved: Unknown$0.09$0.18View docsNot providedvisa · mastercardunknown
ProAI APIunknownLive API pricePlatform pageObserved: Unknown$0.1103$0.2207View docsNot providedalipay · wechatunknown
CrazyRouterunknownLive API pricePlatform pageObserved: Unknown$0.126$0.252View docsNot providedunknown
LingYa APIunknownLive API pricePlatform pageObserved: Unknown$0.1379$0.2759View docsNot providedalipay · wechatunknown
OhMyGPTunknownLive API pricePlatform pageObserved: Unknown$0.1379$0.2759View docsNot providedalipay · wechatunknown
PackyCodeunknownLive API pricePlatform pageObserved: Unknown$0.1379$0.2759View docsNot providedalipay · wechatunknown
PoloAPIunknownLive API pricePlatform pageObserved: Unknown$0.1379$0.2759View docsNot providedalipay · wechatunknown
UiUiAPIunknownLive API pricePlatform pageObserved: Unknown$0.1379$0.2759View docsNot providedalipay · wechat · visa · mastercardunknown
Fireworks AIunknownLive API pricePlatform pageObserved: Unknown$0.14$0.28View docsNot providedvisa · mastercardunknown
OpenRouterunknownLive API pricePlatform pageObserved: Unknown$0.14$0.28View docsNot providedvisa · mastercard · cryptounknown
SiliconFlowunknownLive API pricePlatform pageObserved: Unknown$0.1476$0.2952View docsNot providedalipay · wechatunknown
YunWu AIunknownLive API pricePlatform pageObserved: Unknown$1.00$2.00View docsNot providedalipay · wechatunknown

Platform-submitted information or domain verification does not mean the price was independently checked and is not a ComputeUnion guarantee of pricing, availability, refunds, compliance, invoices, or service quality.

DeepSeek V4 Flash price history

Observed input-price changes across providers over the last 30/90 days.

Unit: $/1M tokens

Estimate monthly cost

The estimate uses the displayed price without assuming discounts or cache hits.

Estimated monthly cost$0.05

FAQ

How is DeepSeek V4 Flash API priced?

Input and output tokens are billed separately. The current recorded official baseline is $0.1400 input / $0.2800 output per 1M tokens.

Which provider is cheapest for DeepSeek V4 Flash?

The lowest recorded input price is $0.0158/1M tokens from RunAPI; verify provider terms and the observation date before production use.

How do official and relay prices differ?

Official and relay channels can differ in price, latency, limits, and service terms; evaluate them separately for production traffic.