Pricing | OpenAI API
For the complete documentation index, see llms.txt. Markdown versions of documentation pages are available by appending `.md` to the page URL.

Docs Guides, concepts, and product docs for CodexUse cases Example workflows and tasks teams can take on with ChatGPT or Codex
Plugins Extend ChatGPT and CodexWorkspace Agents Trigger published ChatGPT workspace agentsCommerce Build commerce flows in ChatGPTAds Publish and measure ads in ChatGPT
Showcase Demo apps to get inspiredBlog Learnings and experiences from developersCookbook Notebook examples for building with OpenAI modelsLearn Docs, videos, and demo apps for building with OpenAICommunity Programs, meetups, and support for builders
Start searching
OverviewModelsAgentsToolsVoice & AudioProductionAPI reference
Search the API docs
Search docs
Suggested
response_format reasoning_effort streaming tools
Primary navigation
API Codex ChatGPT Docs Use cases Training Resources Resources
Search docs
Suggested
response_format reasoning_effort streaming tools
Overview Models Agents Tools Voice & Audio Production API reference
Docs section Models
- Home
Get started
Core concepts
- Webhooks
SDKs and CLI
Resources
Legacy APIs
* Agent Builder
- Overview
* Evals
- Graders
* Fine-tuning
- Direct preference optimization
* Assistants API
Choose a model
- Pricing
Text and code
Prompting
- Overview
Reasoning
Images and video
Realtime and audio
- Overview
Specialized models
- Overview
Agents SDK
- Integrations and observability
ChatKit
- Overview
- Widgets
- Actions
- Overview
Search and retrieval
Connect tools and data
Build tool workflows
- Skills
Computer and code
- Shell
Media
- Overview
Get started
Audio
Connection methods
- WebRTC
- SIP
Sessions and operations
- Webhooks and server-side controls
Go live
Performance and quality
Cost and throughput
- Batch
Safety and governance
Infrastructure and access
- Overview
- Model, tool, and data controls
* Workload identity federation
- AWS
- SPIFFE
Operations
Docs section Docs
Plugins Workspace Agents Commerce Ads
Docs section Select...
- Home
Core concepts
- Skills
Plan
Build
- Add UI to your MCP server (optional)
- Examples
Test and publish
- Connect and test your plugin
Conversion specs
Guides
Resources
- MCP server review requirements
- Home
Get started
- Trigger workspace agent runs
- Authenticate with Workspace Agent access tokens
- Home
Guides
File Upload
- Overview
- Products
API
- Overview
- Feeds
- Products
Measurement
Advertiser API
- Overview
- Bulk API
- Conversion-Optimized Campaigns
API Reference
- Ads
- Insights
- Files
Overview Features Configuration Developers Security Administration Use Cases Resources
Docs section Overview
- Home
Get started
Foundations
Explore
- Models
- Pricing
- Glossary
Available on
- Remote
Releases
- Overview
Workflows
- Sites
- Pets
Capabilities
- Browser
- Voice
- Plugins
- Appshots
Reference
- Commands
- Settings
- Overview
Customization
- Overview
- Memories
Config file
Agent configuration
- Speed
- Rules
Extend ChatGPT and Codex
- MCP
Linux
Windows
- WSL
- Overview
Development workflows
Extend and automate
- Hooks
Environments
- Modes
Build with Codex
Third-party integrations
- GitHub
- Slack
- Linear
Reference
- Overview
Permissions
- Profiles
Codex Security
- Overview
* Codex Security plugin
* Codex Security CLI
- FAQ
* Codex Security cloud
- Setup
- FAQ
Cyber safety
- Overview
Getting started
- ChatGPT Work: usage and cost
Identity and authentication
Workspace access, policy, and models
- Roles and workspace permissions
- Workspace model availability
Plugin and connector controls
Usage, governance, and compliance
- Compliance API and audit events
Deployment and model providers
- Home
- Videos
- Showcase
Community
- Meetups
Blog
- Home
- Videos
- Showcase
Community
- Meetups
Blog
Showcase Blog Cookbook Learn Community
Docs section Select...
Recent
- Meet Rosalind Workbench: Empowering every scientist to be their own research team
- Automating repetitive work at OpenAI with Codex
- Meet the winners of OpenAI Build Week
- Scaling cyber defenders with Daybreak
- Codex as a platform: build on the open agent harness
Topics
- General
- API
- Apps SDK
- Audio
- Codex
- Home
Topics
- Agents
- Evals
- Text
- ChatGPT
- Codex
- gpt-oss
Contribute
- Home
- Docs MCP
Categories
- Videos
Topics
- Agents
- Codex
- Evals
- gpt-oss
- Scaling
- Tools
Programs
Events
- Meetups
Spaces
- Discord
- X
Choose a model
- Pricing
Text and code
Prompting
- Overview
Reasoning
Images and video
Realtime and audio
- Overview
Specialized models
- Tools
Copy Page
Pricing
Copy Page
Flagship models
Our latest models
Prices per 1M tokens.
Standard Batch Flex Fast mode
Standard
| | Short context | Long context | | --- | --- | --- | | Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output | | gpt-5.6-sol | $4.00 | $0.40 | $5.00 | $20.00 | $8.00 | $0.80 | $10.00 | $30.00 | | gpt-5.6-terra | $2.00 | $0.20 | $2.50 | $12.00 | $4.00 | $0.40 | $5.00 | $18.00 | | gpt-5.6-luna | $0.20 | $0.02 | $0.25 | $1.20 | $0.40 | $0.04 | $0.50 | $1.80 |
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details. OpenAI models in Amazon Bedrock are billed through AWS and may differ from direct OpenAI pricing.
Priority processing was renamed Fast mode on July 30, 2026. You can use either `service_tier: "priority"` or `service_tier: "fast"` in your API requests. Learn more about Fast mode.
GPT-5.6 Sol’s promotional pricing is available at least through November 21, 2026.
All models
Batch
| | Short context | Long context | | --- | --- | --- | | Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output | | gpt-5.6-sol | $2.00 | $0.20 | $2.50 | $10.00 | $4.00 | $0.40 | $5.00 | $15.00 | | gpt-5.6-terra | $1.00 | $0.10 | $1.25 | $6.00 | $2.00 | $0.20 | $2.50 | $9.00 | | gpt-5.6-luna | $0.10 | $0.01 | $0.125 | $0.60 | $0.20 | $0.02 | $0.25 | $0.90 |
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.
All models
Flex
| | Short context | Long context | | --- | --- | --- | | Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output | | gpt-5.6-sol | $2.00 | $0.20 | $2.50 | $10.00 | $4.00 | $0.40 | $5.00 | $15.00 | | gpt-5.6-terra | $1.00 | $0.10 | $1.25 | $6.00 | $2.00 | $0.20 | $2.50 | $9.00 | | gpt-5.6-luna | $0.10 | $0.01 | $0.125 | $0.60 | $0.20 | $0.02 | $0.25 | $0.90 |
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.
All models
Fast mode
| | Short context | Long context | | --- | --- | --- | | Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output | | gpt-5.6-sol | $8.00 | $0.80 | $10.00 | $40.00 | $16.00 | $1.60 | $20.00 | $60.00 | | gpt-5.6-terra | $4.00 | $0.40 | $5.00 | $24.00 | $8.00 | $0.80 | $10.00 | $36.00 | | gpt-5.6-luna | $0.40 | $0.04 | $0.50 | $2.40 | $0.80 | $0.08 | $1.00 | $3.60 |
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.
All models
Cyber models
Our latest Daybreak models.
Prices per 1M tokens.
| | Short context | Long context | | --- | --- | --- | | Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output | | gpt-5.6-sol | $4.00 | $0.40 | $5.00 | $20.00 | $8.00 | $0.80 | $10.00 | $30.00 | | gpt-5.6-cyber | $12.50 | $1.25 | $15.625 | $75.00 | - | - | - | - |
All models
`gpt-daybreak-blue-latest` and `gpt-daybreak-red-latest` are aliases that currently point to `gpt-5.6-sol` and `gpt-5.6-cyber`, respectively. As new models are released through the Daybreak program, these aliases will be updated to point to the latest models, with pricing adjusted to match each underlying model.
Multimodal models
To estimate vision model input costs, use the image input cost calculator.
Realtime and audio generation models
Prices per 1M tokens unless noted.
| Model | Modality | Input | Cached input | Output / cost | | --- | --- | --- | --- | --- | | gpt-realtime-2.1 | Audio | $32.00 | $0.40 | $64.00 | | Text | $4.00 | $0.40 | $24.00 | | Image | $5.00 | $0.50 | - | | gpt-realtime-2.1-mini | Audio | $10.00 | $0.30 | $20.00 | | Text | $0.60 | $0.06 | $2.40 | | Image | $0.80 | $0.08 | - |
All models
Image generation models
Prices per 1M tokens.
Standard Batch
Standard
For image generation cost estimates, use the calculator in the image generation guide.
| Model | Modality | Input | Cached input | Output | | --- | --- | --- | --- | --- | | gpt-image-2 | Image | $8.00 | $2.00 | $30.00 | | Text | $5.00 | $1.25 | - |
All models
Batch
For image generation cost estimates, use the calculator in the image generation guide.
| Model | Modality | Input | Cached input | Output | | --- | --- | --- | --- | --- | | gpt-image-2 | Image | $4.00 | $1.00 | $15.00 | | Text | $2.50 | $0.625 | - |
All models
Video generation models
Prices per second.
Standard Batch
Standard
| Model | Size | Portrait | Landscape | Price per second | | --- | --- | --- | --- | --- | | sora-2 | 720p | 720x1280 | 1280x720 | $0.10 | | sora-2-pro | 720p | 720x1280 | 1280x720 | $0.30 | | 1024p | 1024x1792 | 1792x1024 | $0.50 | | 1080p | 1080x1920 | 1920x1080 | $0.70 |
Batch
| Model | Size | Portrait | Landscape | Price per second | | --- | --- | --- | --- | --- | | sora-2 | 720p | 720x1280 | 1280x720 | $0.05 | | sora-2-pro | 720p | 720x1280 | 1280x720 | $0.15 | | 1024p | 1024x1792 | 1792x1024 | $0.25 | | 1080p | 1080x1920 | 1920x1080 | $0.35 |
Transcription models
Prices per 1M tokens unless noted.
| Model | Use case | Input | Output | Estimated cost | | --- | --- | --- | --- | --- | | gpt-realtime-translate | Live translation | - | - | $0.034 / minute | | gpt-live-transcribe | Live transcription | - | - | $0.017 / minute | | gpt-realtime-whisper | Live transcription | - | - | $0.017 / minute | | gpt-transcribe | Transcription | - | - | $0.0045 / minute | | gpt-4o-transcribe | Transcription | $2.50 | $10.00 | $0.006 / minute | | gpt-4o-mini-transcribe | Transcription | $1.25 | $5.00 | $0.003 / minute |
All models
Tools
| Tool | Details | Pricing | | --- | --- | --- | | Web search | Web search (all models) | $10.00 / 1k calls + Search content tokens billed at model rates. | | Image Web search (all models) | $10.00 / 1k calls + Search content tokens billed at model rates. | | Web search preview (reasoning models, including `gpt-5`, `o-series`) | $10.00 / 1k calls + Search content tokens billed at model rates. | | Web search preview (non-reasoning models) | $25.00 / 1k calls + Search content tokens are free. | | Containers | Hosted Shell and Code Interpreter | 1 GB $0.03, 4 GB $0.12, 16 GB $0.48, 64 GB $1.92 per 20-minute session per container. | | File search | Storage | $0.10 / GB per day (1 GB free) | | Tool call | $2.50 / 1k calls | | Agent Kit | ChatKit file and image upload storage | $0.10 / GB-day after 1 GB free per account per month |
Tokens used for built-in tools are billed at the chosen model's per-token rates. GB refers to binary gigabytes (also known as gibibytes), where 1 GB is 2^30 bytes. Web search content tokens are tokens retrieved from the search index and fed to the model alongside your prompt to generate an answer. For `gpt-4o-mini` and `gpt-4.1-mini` with the non-preview web search tool, search content tokens are billed as a fixed block of 8,000 input tokens per call. File search tool call pricing applies to the Responses API only. Container pricing includes Hosted Shell and Code Interpreter. Eligible container sessions will be billed by the minute, with a 5-minute minimum per session. Responses API, Chat Completions API, Realtime API, Batch API, and Assistants API are not priced separately. Tokens are billed at the chosen model's input and output rates.
Specialized models
Prices per 1M tokens.
Standard Fast mode
Standard
| Category | Model | Input | Cached input | Output | | --- | --- | --- | --- | --- | | ChatGPT | chat-latest | $5.00 | $0.50 | $30.00 | | Codex | gpt-5.3-codex | $1.75 | $0.175 | $14.00 |
All models
Fast mode
| Category | Model | Input | Cached input | Output | | --- | --- | --- | --- | --- | | Codex | gpt-5.3-codex | $3.50 | $0.35 | $28.00 |
Finetuning
Prices per 1M tokens.
OpenAI is winding down the fine-tuning platform. The platform is no longer accessible to new users, but existing users of the fine-tuning platform will be able to create training jobs for the coming months.
All fine-tuned models will remain available for inference until their base models are deprecated. The full timeline is here.
Standard Batch
Standard
| Model | Training | Input | Cached input | Output | | --- | --- | --- | --- | --- | | o4-mini-2025-04-16 | $100.00 / hour | $4.00 | $1.00 | $16.00 | | o4-mini-2025-04-16 with data sharing | $100.00 / hour | $2.00 | $0.50 | $8.00 |
All models
Batch
| Model | Training | Input | Cached input | Output | | --- | --- | --- | --- | --- | | o4-mini-2025-04-16 | $100.00 / hour | $2.00 | $0.50 | $8.00 | | o4-mini-2025-04-16 with data sharing | $100.00 / hour | $1.00 | $0.25 | $4.00 |
All models
Tokens used for model grading in reinforcement fine-tuning are billed at that model's per-token rate. Inference discounts are available if you enable data sharing when creating the fine-tune job. Learn more.
Ask AI
Docs agent
Loading docs agent...