AI Model Cost Calculator
Compare editable AI API price presets and estimate token costs, cache savings, monthly usage, gateway margin, and sustainable pricing.
Editable usage, rate, and operating assumptions
Scenario totals and unit economics
Compare official API model price structures
AI Model Cost & Gateway Pricing Calculator
Estimate official API token costs, cache savings, delivery costs, and sustainable pricing for a compliant AI product.
Selecting a preset updates the editable token rates below.
OpenAI GPT-5.6 Sol · gpt-5.6-sol · GA · verified 2026-08-03
Long-context pricing applies to the whole request when average input tokens are more than 272,000. Crossing the boundary reloads the official tier; other edits stay manual.
Official standard and batch token rates; unpublished price components remain absent. Rates above 272k prompt tokens use the long-context tier.
| Item | Basis | Value |
|---|---|---|
| Uncached input | 273M tokens at $5.00/1M | $1,365 |
| Cached input | 147M tokens at $0.50/1M | $73.50 |
| Output tokens | 210M tokens at $30.00/1M | $6,300 |
| Tool and media add-ons | 300,000 requests | $0.00 |
| Retry cost | 2% of base upstream cost | $155 |
| Total direct cost | Upstream API plus monthly platform cost | $8,243 |
| Recommended price | Per 1M blended tokens | $23.58 |
| Planned net profit | 4.5% revenue deductions | -$3,468.27 |
Reference presets verified 2026-08-03. Prices change frequently. Always verify the linked official provider page before making billing decisions.
Cache writes, hourly cache storage, regional processing, enterprise terms, taxes, and other provider-specific charges are disclosed where known but are not silently added. Enter any applicable amount through the editable rates or manual add-on fields.
Pricing methodologyMachine-readable JSON
For official API pricing, BYOK products, team planning, and compliant gateway economics. Not for account sharing, credential resale, or proxy access.
Frequently asked questions
Practical answers for applying this calculator to a production API billing or usage plan.
Are the AI model prices kept up to date automatically?
No. The calculator starts with a small set of reference presets, but every price field is editable. Providers change pricing, model availability, cache rules, and tool fees, so verify the current official provider page before using the result in a billing decision.
How does cache hit rate affect AI API cost?
The calculator treats the selected percentage of input tokens as cached and prices those tokens at the cached-input rate. A higher cache hit rate lowers cost only when the provider offers a discounted cache-read price and your requests reuse eligible prompt content.
Does the recommended revenue include payment fees and refunds?
Yes. Payment fees, expected refunds, and bad debt are modeled as percentages of revenue. They are included with target margin when calculating the monthly revenue needed to cover direct delivery and platform costs.
Can I use this for account sharing or subscription proxy access?
No. It is designed for official API pricing, BYOK products, legitimate team planning, and compliant AI gateway economics. It is not for account sharing, credential resale, or proxy access.
Related API billing guides
Read the concepts behind the calculator and adapt the examples to your own API pricing model.
How to Price an API: Cost, Margin and Billing Models
Follow a step-by-step method to choose a billable unit, calculate delivery cost, set margin, design free tiers, and publish sustainable paid API plans.
Read guideAPI Fees & Charges Explained: Rates and Invoice Examples
Understand API fees and charges by request, token, credit, and plan, including free tiers, overage, rounding, and worked invoice lines.
Read guideAPI Billing Models: Usage-Based, Subscription, Credits and BYOK
Compare API billing models for developer products, including usage-based billing, subscriptions, credits, prepaid wallets, and bring-your-own-key.
Read guide