AlibabaLLMactive

QwQ 32B API Pricing

Compare prices from 3 providers, including 0 official. 2 price sources are collected automatically.

Data snapshot:

Official input / 1M
Lowest input / 1M$0.30
Maximum input gapNot a service-quality measure
Comparable providers32 live

Task fit and model switching guide

Derived from official positioning and current ComputeUnion provider and pricing data.

Official positioningOfficially positioned for a specialized workload.

Discount vs official API

Input-price difference for the same model; negative values are above official price.

Official baseline unavailable; discount is not calculated.

Price composition

Input and output remain paired within each provider quote.

Capability radar

Task-fit levels derived from official positioning, not an independent benchmark.

Channel unit price and evidence map

X-axis is the provider quote per 1M input tokens; Y-axis is price-evidence strength. No usage volume is assumed.

External capability comparison

Not enough verified related-model data is available for comparison.

Sources and verification method
  • Prices: current official and provider quotes in ComputeUnion, calculated within each provider.
  • Capability: official task positioning plus safely matched Artificial Analysis metrics; neither is presented as a ComputeUnion benchmark.
  • Evidence: official, live API, platform-submitted, and unreviewed quotes remain visibly separated.

Best-fit tasks

  • Best suited to complex reasoning.
  • Best suited to coding.
  • Suitable for research.

When to switch models

  • Switch to Qwen3.7 Max when quality and complex reasoning take priority.
  • Compare Qwen3.6 35B A3B when lower cost with strong capability matters.
  • Compare Qwen3.5 9B for batch, repeatable, or cost-sensitive workloads.

ComputeUnion market view

  • ComputeUnion currently tracks 3 providers, including 0 official and 2 live price sources.
  • 1 quote has not been independently checked; verify provider terms before payment.

Task guidance is generated by fixed rules from official positioning and ComputeUnion market data; it is not an independent benchmark or guarantee.

QwQ 32B provider prices and multipliers

No complete verified official input/output baseline is currently recorded, so provider quotes are shown without multipliers.

Providers: 3From, within billing group: $0.2759Lowest observed: $0.2759Lowest input multiplier:
Official input baselineOfficial output baselinePrice multipliers do not measure stability, rate limits, refunds, or service quality.
QwQ 32B provider price and evidence comparison
Provider / operatorInput / 1M tokensOutput / 1M tokensvs officialPrice evidencePublic operating historyAPI docsRefund policyPayment / invoice
CompshareOperator type not confirmed$0.2759$0.8276
In Out
Live API pricePlatform pageObserved: UnknownView docsNot providedAlipay · WeChat Pay · bankcardInvoice support not confirmed
Cloudflare Workers AIOperator type not confirmed$0.66$1.00
In Out
Live API pricePlatform pageObserved: UnknownView docsNot providedVisa · MastercardInvoice support not confirmed
Novita AIOperator type not confirmed$0.30$0.30
In Out
UnreviewedPlatform pageObserved: UnknownView docsNot providedVisa · MastercardInvoice support not confirmed

Platform-submitted information or domain verification does not mean the price was independently checked and is not a ComputeUnion guarantee of pricing, availability, refunds, compliance, invoices, or service quality.

QwQ 32B price history

Observed input-price changes across providers over the last 30/90 days.

Unit: $/1M tokens

FAQ

How is QwQ 32B API priced?

Input and output tokens are billed separately. A complete official input/output baseline is not currently recorded, so provider quotes are normalized per million tokens.

Which provider is cheapest for QwQ 32B?

The lowest recorded input price is $0.30/1M tokens from Novita AI; verify provider terms and the observation date before production use.

How do official and relay prices differ?

Official and relay channels can differ in price, latency, limits, and service terms; evaluate them separately for production traffic.