GPT-4.1 API Pricing
16 available providers and 15 live price sources tracked, including 1 official provider.
Data snapshot:
Task fit and model switching guide
Derived from official positioning and current ComputeUnion provider and pricing data.
Discount vs official API
Input-price difference for the same model; negative values are above official price.
Price composition
Input and output remain paired within each provider quote.
Capability radar
Task-fit levels derived from official positioning, not an independent benchmark.
Channel unit price and evidence map
X-axis is the provider quote per 1M input tokens; Y-axis is price-evidence strength. No usage volume is assumed.
External capability comparison
Not enough verified related-model data is available for comparison.
Choose a related model for GPT-4.1
Compare price, capability, and provider coverage; select a model name for its detail page.
| Model | Official role | Official input/output | External score | Providers |
|---|---|---|---|---|
| GPT-4o | Not classified | $2.5 / $10 | — | 15 |
| Claude Sonnet 4.6 | Balanced workhorse | $3 / $15 | 34 | 26 |
| Gemini 2.5 Pro | Frontier flagship | $1.25 / $10 | 26 | 20 |
- Prices: current official and provider quotes in ComputeUnion, calculated within each provider.
- Capability: official task positioning plus safely matched Artificial Analysis metrics; neither is presented as a ComputeUnion benchmark.
- Evidence: official, live API, platform-submitted, and unreviewed quotes remain visibly separated.
Best-fit tasks
- Best suited to coding.
- Suitable for general production workloads.
- Suitable for structured transformation and summaries.
When to switch models
- Switch to GPT-5.5 when quality and complex reasoning take priority.
- Compare GPT-4.1 mini when lower cost with strong capability matters.
ComputeUnion market view
- ComputeUnion currently tracks 16 providers, including 1 official and 15 live price sources.
- The lowest recorded input price is 95.9% below the official input price; a lower price does not imply better reliability, limits, or refund terms.
Task guidance is generated by fixed rules from official positioning and ComputeUnion market data; it is not an independent benchmark or guarantee.
GPT-4.1 provider prices and multipliers
Input and output multipliers are each calculated as provider price ÷ official baseline; 1.00× equals the official price.
| Provider / operator | Input / 1M tokens | Output / 1M tokens | vs official | Price evidence | Public operating history | API docs | Refund policy | Payment / invoice |
|---|---|---|---|---|---|---|---|---|
| Azure OpenAIunknown | $2.00 | $8.00 | In 1.00×Out 1.00× | Official price sourcePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | —unknown |
| Qingfengge APIVerified platformteam | $0.1379 | $0.5517 | In 0.069×Out 0.069× | Live API pricePlatform page ↗Observed: | 3 years 7 monthsDomain verified | View docs ↗ | Refund policy not public | 支付宝、微信、对公打款china tax invoice |
| TreeRouterunknown | $0.0828 | $0.331 | In 0.0414×Out 0.0414× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechatunknown |
| RunAPIunknown | $0.1103 | $0.4414 | In 0.0552×Out 0.0552× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechatunknown |
| EasyRouterunknown | $0.2759 | $1.1034 | In 0.14×Out 0.14× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechatunknown |
| LaoZhang APIunknown | $0.2759 | $1.1034 | In 0.14×Out 0.14× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechatunknown |
| PackyCodeunknown | $0.2759 | $0.5517 | In 0.14×Out 0.069× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechatunknown |
| PoloAPIunknown | $0.2759 | $1.1034 | In 0.14×Out 0.14× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechatunknown |
| UiUiAPIunknown | $0.2759 | $1.1034 | In 0.14×Out 0.14× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechat · visa · mastercardunknown |
| ProAI APIunknown | $0.2869 | $1.1476 | In 0.14×Out 0.14× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechatunknown |
| CrazyRouterunknown | $1.30 | $5.20 | In 0.65×Out 0.65× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | —unknown |
| 302.AIunknown | $2.00 | $8.00 | In 1.00×Out 1.00× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · visa · mastercard · amex · unionpayunknown |
| OpenRouterunknown | $2.00 | $8.00 | In 1.00×Out 1.00× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | visa · mastercard · cryptounknown |
| YunWu AIunknown | $2.00 | $8.00 | In 1.00×Out 1.00× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechatunknown |
| OhMyGPTunknown | $2.20 | $8.80 | In 1.10×Out 1.10× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechatunknown |
| Poixe AIunknown | $24.00 | $96.00 | In 12.0×Out 12.0× | Live API pricePlatform page ↗Observed: | —Unknown | View docs ↗ | Not provided | alipay · wechat · visa · mastercardunknown |
Platform-submitted information or domain verification does not mean the price was independently checked and is not a ComputeUnion guarantee of pricing, availability, refunds, compliance, invoices, or service quality.
GPT-4.1 price history
Observed input-price changes across providers over the last 30/90 days.
Unit: $/1M tokens
FAQ
How is GPT-4.1 API priced?
Input and output tokens are billed separately. The current recorded official baseline is $2.0000 input / $8.0000 output per 1M tokens.
Which provider is cheapest for GPT-4.1?
The lowest recorded input price is $0.0828/1M tokens from TreeRouter; verify provider terms and the observation date before production use.
How do official and relay prices differ?
Official and relay channels can differ in price, latency, limits, and service terms; evaluate them separately for production traffic.