OpenAI API Token Pricing: Costs per 1M Tokens
ai-costs

OpenAI API Token Costs Explained: Prices Per Model and Per Workflow

By Paige Gilmore, Founder, NetLift· Published July 28, 2026· Updated July 28, 2026

OpenAI API pricing is based on token volume, with input costs ranging from $0.20 to $30 per 1M tokens. Output costs are significantly higher, reaching up to $180 per 1M tokens for premium models.

14-day free trial. No credit card required.

OpenAI charges for API access using a consumption-based model measured in tokens. Costs depend heavily on the specific model and the type of data being processed, with input tokens generally priced much lower than output tokens. For instance, the entry-level gpt-5.4-nano costs $0.20 per 1M input tokens, while the high-performance gpt-5.5-pro costs $30 per 1M input tokens.

Output costs scale with model complexity. While gpt-5.6-luna provides a cost-effective output rate of $6 per 1M tokens, the gpt-5.5-pro model reaches $180 per 1M tokens. Specialized tasks like audio processing or image input carry their own specific rates, such as $8 per 1M tokens for image data.

Official per-token prices

Model / rate Price Unit
OpenAI API — gpt-5.4-nano (Standard Short context) input $0.2 per 1M input tokens
OpenAI API — gpt-5.4-mini (Standard Short context) input $0.75 per 1M input tokens
OpenAI API — gpt-5.6-luna (Standard Short context) input $1 per 1M input tokens
OpenAI API — gpt-5.6-terra (Standard Short context) input $2.50 per 1M input tokens
OpenAI API — gpt-5.4-mini (Standard Short context) output $4.50 per 1M output tokens
OpenAI API — gpt-5.6-sol (Standard Short context) input $5 per 1M input tokens
OpenAI API — gpt-5.5 (Standard Short context) input $5 per 1M input tokens
OpenAI API — o3-deep-research input $5 per 1M input tokens
OpenAI API — gpt-5.6-luna (Standard Short context) output $6 per 1M output tokens
OpenAI API — gpt-5.6-sol (Standard Long context) input $10 per 1M input tokens
OpenAI API — gpt-5.6-terra (Standard Short context) output $15 per 1M output tokens
OpenAI API — o3-deep-research output $20 per 1M output tokens
OpenAI API — gpt-5.6-sol (Standard Short context) output $30 per 1M output tokens
OpenAI API — gpt-5.5 (Standard Short context) output $30 per 1M output tokens
OpenAI API — gpt-5.5-pro (Standard Short context) input $30 per 1M input tokens
OpenAI API — gpt-realtime-2.1 Audio input $32 per 1M input tokens
OpenAI API — gpt-5.6-sol (Standard Long context) output $45 per 1M output tokens
OpenAI API — gpt-realtime-2.1 Audio output $64 per 1M output tokens
OpenAI API — gpt-5.5-pro (Standard Short context) output $180 per 1M output tokens
chat-latest — ChatGPT standard $5 per 1M input tokens
gpt-5.3-codex — Codex standard input $1.75 per 1M input tokens
gpt-5.4 — Standard Short context Input $2.50 per 1M input tokens
gpt-5.4 — Standard Short context Output $15 per 1M output tokens
gpt-5.4-mini — Standard Short context Input $0.75 per 1M input tokens
gpt-5.4-mini — Standard Short context Output $4.50 per 1M output tokens
gpt-5.4-nano — Standard Short context Input $0.2 per 1M input tokens
gpt-5.4-nano — Standard Short context Output $1.25 per 1M output tokens
gpt-5.5 — Standard Short context Input $5 per 1M input tokens
gpt-5.5 — Standard Short context Output $30 per 1M output tokens
gpt-5.5-pro — Standard Short context Input $30 per 1M input tokens
gpt-5.5-pro — Standard Short context Output $180 per 1M output tokens
gpt-5.6-luna — Standard Short context Input $1 per 1M input tokens
gpt-5.6-luna — Standard Short context Output $6 per 1M output tokens
gpt-5.6-sol — Standard Short context Cached input $0.5 per 1M input tokens
gpt-5.6-sol — Standard Short context Input $5 per 1M input tokens
gpt-5.6-sol — Standard Long context Input $10 per 1M input tokens
gpt-5.6-sol — Standard Short context Output $30 per 1M output tokens
gpt-5.6-sol — Standard Long context Output $45 per 1M output tokens
gpt-5.6-terra — Standard Short context Input $2.50 per 1M input tokens
gpt-5.6-terra — Standard Short context Output $15 per 1M output tokens

14-day free trial. No credit card required.

How do context length and caching affect the bill?

Price points change based on how the model handles data. The gpt-5.6-sol model illustrates this variability: standard short context input costs $5 per 1M tokens, but this drops to $0.50 if the input is cached. Conversely, if the task requires long context processing, the rate for the same model increases to $10 per 1M input tokens and $45 per 1M output tokens.

What are the costs for audio and specialized research?

Multimedia and deep reasoning tasks use different pricing tiers. Real-time audio via gpt-realtime-2.1 is one of the more expensive services at $32 for input and $64 for output per 1M tokens. For high-volume research tasks, o3-deep-research offers a batch input rate of $5 per 1M tokens, providing a structured cost for intensive data processing.

To determine if these token costs are a sound investment, organizations must measure the cost per task against the human time saved. A model like gpt-5.6-terra, costing $2.50 per 1M input tokens, may be more capital-efficient than a pro-tier model if the output quality meets the required threshold. NetLift allows finance teams to track this unit economics by comparing API spend directly against labor hours reclaimed, ensuring AI adoption delivers a measurable net return.

14-day free trial. No credit card required.

Sources

All pricing on this page comes from official vendor pages:

Frequently Asked Questions

Ready to see your AI return?

14-day free trial. No credit card required.

About the author

Paige Gilmore · Founder, NetLift

Paige Gilmore is the founder of NetLift, the AI Value Management platform that helps organisations measure the cost, savings and return of AI adoption.

Paige Gilmore on LinkedIn

Keep reading