p.enthalabs

Pricing | OpenAI API

For the complete documentation index, see llms.txt. Markdown versions of documentation pages are available by appending `.md` to the page URL.

![Image 1: OpenAI DevelopersChatGPT](https://developers.openai.com/)

Home

API

Codex

Docs Guides, concepts, and product docs for CodexUse cases Example workflows and tasks teams can take on with ChatGPT or Codex

Docs

Use cases

Training

Resources

ChatGPT

Plugins Extend ChatGPT and CodexWorkspace Agents Trigger published ChatGPT workspace agentsCommerce Build commerce flows in ChatGPTAds Publish and measure ads in ChatGPT

Resources

Showcase Demo apps to get inspiredBlog Learnings and experiences from developersCookbook Notebook examples for building with OpenAI modelsLearn Docs, videos, and demo apps for building with OpenAICommunity Programs, meetups, and support for builders

Start searching

API Dashboard

Try ChatGPT

OverviewModelsAgentsToolsVoice & AudioProductionAPI reference

Search the API docs

Search docs

Suggested

response_format reasoning_effort streaming tools

Primary navigation

API Codex ChatGPT Docs Use cases Training Resources Resources

Search docs

Suggested

response_format reasoning_effort streaming tools

Overview Models Agents Tools Voice & Audio Production API reference

Docs section Models

- Home

Get started

- Quickstart

- Using GPT-5.6

- Key concepts

Core concepts

- Responses API

- Conversation state

- Background mode

- Streaming

- WebSocket mode

- Multi-agent

- Webhooks

- File inputs

- Compaction

- Counting tokens

SDKs and CLI

- OpenAI SDK

- OpenAI CLI

Resources

- Changelog

- Deprecations

- Supported countries

- OpenAI Crawlers

- Terms and policies

Legacy APIs

* Agent Builder

- Overview

- Migration guide

- Node reference

- Safety in building agents

* Evals

- Getting started

- Working with evals

- Prompt optimizer

- External models

- Best practices

- Graders

* Fine-tuning

- Optimization cycle

- Supervised fine-tuning

- Vision fine-tuning

- Direct preference optimization

- Reinforcement fine-tuning

- RFT use cases

- Best practices

* Assistants API

- Migration guide

- Model catalog

Choose a model

- Pricing

- Model selection

Text and code

- Text generation

- Code generation

- Structured output

Prompting

- Overview

- Prompt engineering

- Citation formatting

- Migration guide

- Prompt generation

- Frontend prompting

Reasoning

- Reasoning models

- Reasoning best practices

Images and video

* Images and vision

- Image input cost calculator

- Image generation

- Video generation

Realtime and audio

- Audio and speech

- Overview

- Voice agents

Specialized models

- Deep research

- Embeddings

- Moderation

- Overview

Agents SDK

- Quickstart

- Agent definitions

- Models and providers

- Running agents

- Sandbox agents

- Orchestration

- Guardrails

- Results and state

- Integrations and observability

- Evaluate agent workflows

ChatKit

- Overview

- Customize

- Widgets

- Actions

- Advanced integrations

- Overview

- Function calling

Search and retrieval

- Web search

- File search

- Retrieval

Connect tools and data

- MCP and Connectors

- Secure MCP Tunnel

Build tool workflows

- Skills

- Tool search

- Programmatic tool calling

Computer and code

- Shell

- Computer use

- Apply Patch

- Local shell

- Code interpreter

Media

- Image generation

- Overview

Get started

- Voice agents

- Live translation

- Realtime prompting guide

Audio

- Audio and speech

- Transcription

- File transcription

- Realtime transcription

- Speech generation

Connection methods

- WebRTC

- WebSocket

- SIP

Sessions and operations

- Managing conversations

- Voice activity detection

- Realtime with tools

- Webhooks and server-side controls

- Managing costs

Go live

- Production best practices

- Deployment checklist

Performance and quality

- Latency optimization

- Predicted Outputs

- Fast mode

- Accuracy optimization

Cost and throughput

- Cost optimization

- Prompt caching

- Batch

- Flex processing

Safety and governance

- Safety best practices

- Red teaming

* Safety checks

- Cybersecurity checks

- Under 18 API Guidance

- CSAM guidance

- Content provenance

- Your data

- Permissions

Infrastructure and access

* Terraform provider

- Overview

- Projects and access

- Service accounts

- Rate limits and spend

- Model, tool, and data controls

- Import and reconciliation

- Private Link

- IP allowlist

- Mutual TLS

* Workload identity federation

- Codex setup

- Federation rules

- Admin API

- X.509 certificates

- Kubernetes

- AWS

- Microsoft Azure

- Google Cloud

- Oracle Cloud Infrastructure

- GitHub Actions

- SPIFFE

- IP egress ranges

- Amazon Bedrock

Operations

- Rate limits

- Spend limits

- Admin APIs

- Error codes

DocsUse cases

Docs section Docs

Plugins Workspace Agents Commerce Ads

Docs section Select...

- Home

- Quickstart

Core concepts

- Plugin architecture

- Skills

- MCP server

Plan

- Brainstorm use cases

- Define tools

Build

- Build an MCP server

- Add UI to your MCP server (optional)

- Authenticate users

- Build skills

- Package your plugin

- Examples

Test and publish

- Connect and test your plugin

- Submit and publish

- Submission error reference

Conversion specs

- Restaurant reservation spec

- Get Quote spec

- Product checkout spec

Guides

- UI guidelines

- Optimize Metadata

- Submit a Claude Code plugin

- Security & Privacy

- Troubleshooting

Resources

- Changelog

- Plugin guidelines

- MCP server review requirements

- Plugin UI reference

- Checkout API reference

- Home

Get started

- Trigger workspace agent runs

- Authenticate with Workspace Agent access tokens

- Home

Guides

- Get started

- Best practices

File Upload

- Overview

- Products

API

- Overview

- Feeds

- Products

- Promotions

- Ads Overview

Measurement

- Measurement Pixel

- Multiple Pixels (Advanced)

- Image Tag

- Conversions API

- Supported Events

Advertiser API

- Overview

- API Partner Setup

- Quickstart

- Bulk API

- Product Feeds

- Delta Feeds API

- Campaign Targeting

- Conversion-Optimized Campaigns

- Custom Audiences

API Reference

- Authentication

- Ad Account

- Campaigns

- Ad Groups

- Ads

- Insights

- Files

- Conversion Setup

Overview Features Configuration Developers Security Administration Use Cases Resources

Docs section Overview

- Home

Get started

- Quickstart

- Use ChatGPT

- Get started with Work

- Import from another agent

Foundations

- Prompting

- Personalize ChatGPT

- Skills & Plugins

- Permissions

Explore

- What's new

- Models

- Pricing

- Glossary

Available on

- ChatGPT desktop app

- Remote

- ChatGPT on the web

- Codex CLI

- Codex IDE extension

- Codex cloud

Releases

- Changelog

- Feature Maturity

- Open Source

- Overview

Workflows

- Projects and chats

- Sites

- Visualizations

- Scheduled tasks

- Long-running work

- Notifications

- Pets

- Codex Micro

Capabilities

- Browser

- Computer use

- Voice

- Plugins

- Web search

- Image generation

- Image inputs

- Appshots

- Browser extension

- Work with files

Reference

- Commands

- Slash commands

- Settings

- Troubleshooting

- Overview

Customization

- Overview

- Memories

- Computer History

Config file

- Config Basics

- Advanced Config

- Config Reference

- Environment Variables

- Sample Config

Agent configuration

- AGENTS.md

- Subagents

- Speed

- Rules

Extend ChatGPT and Codex

- Record & Replay

- MCP

Linux

- Desktop app

Windows

- Desktop app

- Windows sandbox

- WSL

- Overview

Development workflows

- Code review

- Integrated terminal

Extend and automate

- Build skills

- Build plugins

- Site tools (WebMCP)

- Hooks

Environments

- Modes

- Local environments

- Cloud environment

- Git worktrees

Build with Codex

- Codex SDK

- App Server

- MCP Server

- GitHub Action

- Non-interactive mode

Third-party integrations

- GitHub

- GitLab (Beta)

- Slack

- Linear

Reference

- CLI customization

- Developer commands

- Developer settings

- Overview

Permissions

- Profiles

- Sandboxing

- Auto-review

- Agent approvals & security

- Internet access

Codex Security

- Overview

* Codex Security plugin

- Quickstart

- Run a security scan

- Run a deep scan

- Review code changes

- Use the Security workbench

- Triage a backlog

- Fix findings

- Propose security hardening

- Write vulnerability reports

- Export and track findings

- Changelog

* Codex Security CLI

- Quickstart

- Run bulk scans

- Run scans in CI

- GitLab CI/CD

- Reference

- FAQ

- TypeScript SDK

* Codex Security cloud

- Setup

- Security Review

- Improving the threat model

- FAQ

Cyber safety

- Models & Trusted Access

- Recommended configuration

- Overview

Getting started

- Admin rollout guide

- ChatGPT Work Overview

- ChatGPT Work cloud security

- ChatGPT Work local security

- ChatGPT Work admin FAQ

- ChatGPT Work: usage and cost

Identity and authentication

- Authentication overview

- Workload identity

- Personal Access Tokens

- Service accounts

Workspace access, policy, and models

- Groups and provisioning

- Roles and workspace permissions

- GPTs and Sharing

- Managed configuration

- Prisma AIRS

- HIPAA configuration

- Workspace model availability

Plugin and connector controls

- Plugin controls

- Plugin management

- Skill controls

Usage, governance, and compliance

- Governance

- Admin plugin

- Workspace analytics

- Analytics API

- Compliance API and audit events

Deployment and model providers

- Manage app updates

- Windows app deployment

- Remote connections

- Amazon Bedrock

- Explore use cases

- Collections

- Home

- Videos

- Showcase

- OpenAI Academy

- Online trainings

Community

- Codex Ambassadors

- Codex for Students

- Codex for Open Source

- Meetups

Blog

- Company blog

- Developer blog

- Explore use cases

- Collections

- Home

- Videos

- Showcase

- OpenAI Academy

- Online trainings

Community

- Codex Ambassadors

- Codex for Students

- Codex for Open Source

- Meetups

Blog

- Company blog

- Developer blog

Showcase Blog Cookbook Learn Community

Docs section Select...

- All posts

Recent

- Meet Rosalind Workbench: Empowering every scientist to be their own research team

- Automating repetitive work at OpenAI with Codex

- Meet the winners of OpenAI Build Week

- Scaling cyber defenders with Daybreak

- Codex as a platform: build on the open agent harness

Topics

- General

- API

- Apps SDK

- Audio

- Codex

- Life sciences

- Home

Topics

- Agents

- Evals

- Multimodal

- Text

- Guardrails

- Optimization

- ChatGPT

- Codex

- gpt-oss

Contribute

- Cookbook on GitHub

- Home

- OpenAI Developers plugin

- Docs MCP

Categories

- Demo apps

- Videos

Topics

- Agents

- Audio & Voice

- Computer Use

- Codex

- Evals

- gpt-oss

- Fine-tuning

- Image generation

- Scaling

- Tools

- Video generation

- Community

Programs

- Codex Ambassadors

- Codex for Students

- Codex for Open Source

- OpenAI for Startups

Events

- Meetups

Spaces

- Developer Forum

- Discord

- Reddit

- X

API Dashboard

Try ChatGPT

- Model catalog

Choose a model

- Pricing

- Model selection

Text and code

- Text generation

- Code generation

- Structured output

Prompting

- Overview

- Prompt engineering

- Citation formatting

- Migration guide

- Prompt generation

- Frontend prompting

Reasoning

- Reasoning models

- Reasoning best practices

Images and video

* Images and vision

- Image input cost calculator

- Image generation

- Video generation

Realtime and audio

- Audio and speech

- Overview

- Voice agents

Specialized models

- Deep research

- Embeddings

- Moderation

- Flagship models

- Cyber models

- Multimodal models

- Tools

- Specialized models

- Finetuning

Copy Page

Pricing

Copy Page

Flagship models

Our latest models

Prices per 1M tokens.

Standard Batch Flex Fast mode

Standard

| | Short context | Long context | | --- | --- | --- | | Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output | | gpt-5.6-sol | $4.00 | $0.40 | $5.00 | $20.00 | $8.00 | $0.80 | $10.00 | $30.00 | | gpt-5.6-terra | $2.00 | $0.20 | $2.50 | $12.00 | $4.00 | $0.40 | $5.00 | $18.00 | | gpt-5.6-luna | $0.20 | $0.02 | $0.25 | $1.20 | $0.40 | $0.04 | $0.50 | $1.80 |

Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details. OpenAI models in Amazon Bedrock are billed through AWS and may differ from direct OpenAI pricing.

Priority processing was renamed Fast mode on July 30, 2026. You can use either `service_tier: "priority"` or `service_tier: "fast"` in your API requests. Learn more about Fast mode.

GPT-5.6 Sol’s promotional pricing is available at least through November 21, 2026.

All models

Batch

| | Short context | Long context | | --- | --- | --- | | Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output | | gpt-5.6-sol | $2.00 | $0.20 | $2.50 | $10.00 | $4.00 | $0.40 | $5.00 | $15.00 | | gpt-5.6-terra | $1.00 | $0.10 | $1.25 | $6.00 | $2.00 | $0.20 | $2.50 | $9.00 | | gpt-5.6-luna | $0.10 | $0.01 | $0.125 | $0.60 | $0.20 | $0.02 | $0.25 | $0.90 |

Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.

All models

Flex

| | Short context | Long context | | --- | --- | --- | | Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output | | gpt-5.6-sol | $2.00 | $0.20 | $2.50 | $10.00 | $4.00 | $0.40 | $5.00 | $15.00 | | gpt-5.6-terra | $1.00 | $0.10 | $1.25 | $6.00 | $2.00 | $0.20 | $2.50 | $9.00 | | gpt-5.6-luna | $0.10 | $0.01 | $0.125 | $0.60 | $0.20 | $0.02 | $0.25 | $0.90 |

Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.

All models

Fast mode

| | Short context | Long context | | --- | --- | --- | | Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output | | gpt-5.6-sol | $8.00 | $0.80 | $10.00 | $40.00 | $16.00 | $1.60 | $20.00 | $60.00 | | gpt-5.6-terra | $4.00 | $0.40 | $5.00 | $24.00 | $8.00 | $0.80 | $10.00 | $36.00 | | gpt-5.6-luna | $0.40 | $0.04 | $0.50 | $2.40 | $0.80 | $0.08 | $1.00 | $3.60 |

Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.

All models

Cyber models

Our latest Daybreak models.

Prices per 1M tokens.

| | Short context | Long context | | --- | --- | --- | | Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output | | gpt-5.6-sol | $4.00 | $0.40 | $5.00 | $20.00 | $8.00 | $0.80 | $10.00 | $30.00 | | gpt-5.6-cyber | $12.50 | $1.25 | $15.625 | $75.00 | - | - | - | - |

All models

`gpt-daybreak-blue-latest` and `gpt-daybreak-red-latest` are aliases that currently point to `gpt-5.6-sol` and `gpt-5.6-cyber`, respectively. As new models are released through the Daybreak program, these aliases will be updated to point to the latest models, with pricing adjusted to match each underlying model.

Multimodal models

To estimate vision model input costs, use the image input cost calculator.

Realtime and audio generation models

Prices per 1M tokens unless noted.

| Model | Modality | Input | Cached input | Output / cost | | --- | --- | --- | --- | --- | | gpt-realtime-2.1 | Audio | $32.00 | $0.40 | $64.00 | | Text | $4.00 | $0.40 | $24.00 | | Image | $5.00 | $0.50 | - | | gpt-realtime-2.1-mini | Audio | $10.00 | $0.30 | $20.00 | | Text | $0.60 | $0.06 | $2.40 | | Image | $0.80 | $0.08 | - |

All models

Image generation models

Prices per 1M tokens.

Standard Batch

Standard

For image generation cost estimates, use the calculator in the image generation guide.

| Model | Modality | Input | Cached input | Output | | --- | --- | --- | --- | --- | | gpt-image-2 | Image | $8.00 | $2.00 | $30.00 | | Text | $5.00 | $1.25 | - |

All models

Batch

For image generation cost estimates, use the calculator in the image generation guide.

| Model | Modality | Input | Cached input | Output | | --- | --- | --- | --- | --- | | gpt-image-2 | Image | $4.00 | $1.00 | $15.00 | | Text | $2.50 | $0.625 | - |

All models

Video generation models

Prices per second.

Standard Batch

Standard

| Model | Size | Portrait | Landscape | Price per second | | --- | --- | --- | --- | --- | | sora-2 | 720p | 720x1280 | 1280x720 | $0.10 | | sora-2-pro | 720p | 720x1280 | 1280x720 | $0.30 | | 1024p | 1024x1792 | 1792x1024 | $0.50 | | 1080p | 1080x1920 | 1920x1080 | $0.70 |

Batch

| Model | Size | Portrait | Landscape | Price per second | | --- | --- | --- | --- | --- | | sora-2 | 720p | 720x1280 | 1280x720 | $0.05 | | sora-2-pro | 720p | 720x1280 | 1280x720 | $0.15 | | 1024p | 1024x1792 | 1792x1024 | $0.25 | | 1080p | 1080x1920 | 1920x1080 | $0.35 |

Transcription models

Prices per 1M tokens unless noted.

| Model | Use case | Input | Output | Estimated cost | | --- | --- | --- | --- | --- | | gpt-realtime-translate | Live translation | - | - | $0.034 / minute | | gpt-live-transcribe | Live transcription | - | - | $0.017 / minute | | gpt-realtime-whisper | Live transcription | - | - | $0.017 / minute | | gpt-transcribe | Transcription | - | - | $0.0045 / minute | | gpt-4o-transcribe | Transcription | $2.50 | $10.00 | $0.006 / minute | | gpt-4o-mini-transcribe | Transcription | $1.25 | $5.00 | $0.003 / minute |

All models

Tools

| Tool | Details | Pricing | | --- | --- | --- | | Web search | Web search (all models) | $10.00 / 1k calls + Search content tokens billed at model rates. | | Image Web search (all models) | $10.00 / 1k calls + Search content tokens billed at model rates. | | Web search preview (reasoning models, including `gpt-5`, `o-series`) | $10.00 / 1k calls + Search content tokens billed at model rates. | | Web search preview (non-reasoning models) | $25.00 / 1k calls + Search content tokens are free. | | Containers | Hosted Shell and Code Interpreter | 1 GB $0.03, 4 GB $0.12, 16 GB $0.48, 64 GB $1.92 per 20-minute session per container. | | File search | Storage | $0.10 / GB per day (1 GB free) | | Tool call | $2.50 / 1k calls | | Agent Kit | ChatKit file and image upload storage | $0.10 / GB-day after 1 GB free per account per month |

Tokens used for built-in tools are billed at the chosen model's per-token rates. GB refers to binary gigabytes (also known as gibibytes), where 1 GB is 2^30 bytes. Web search content tokens are tokens retrieved from the search index and fed to the model alongside your prompt to generate an answer. For `gpt-4o-mini` and `gpt-4.1-mini` with the non-preview web search tool, search content tokens are billed as a fixed block of 8,000 input tokens per call. File search tool call pricing applies to the Responses API only. Container pricing includes Hosted Shell and Code Interpreter. Eligible container sessions will be billed by the minute, with a 5-minute minimum per session. Responses API, Chat Completions API, Realtime API, Batch API, and Assistants API are not priced separately. Tokens are billed at the chosen model's input and output rates.

Specialized models

Prices per 1M tokens.

Standard Fast mode

Standard

| Category | Model | Input | Cached input | Output | | --- | --- | --- | --- | --- | | ChatGPT | chat-latest | $5.00 | $0.50 | $30.00 | | Codex | gpt-5.3-codex | $1.75 | $0.175 | $14.00 |

All models

Fast mode

| Category | Model | Input | Cached input | Output | | --- | --- | --- | --- | --- | | Codex | gpt-5.3-codex | $3.50 | $0.35 | $28.00 |

Finetuning

Prices per 1M tokens.

OpenAI is winding down the fine-tuning platform. The platform is no longer accessible to new users, but existing users of the fine-tuning platform will be able to create training jobs for the coming months.

All fine-tuned models will remain available for inference until their base models are deprecated. The full timeline is here.

Standard Batch

Standard

| Model | Training | Input | Cached input | Output | | --- | --- | --- | --- | --- | | o4-mini-2025-04-16 | $100.00 / hour | $4.00 | $1.00 | $16.00 | | o4-mini-2025-04-16 with data sharing | $100.00 / hour | $2.00 | $0.50 | $8.00 |

All models

Batch

| Model | Training | Input | Cached input | Output | | --- | --- | --- | --- | --- | | o4-mini-2025-04-16 | $100.00 / hour | $2.00 | $0.50 | $8.00 | | o4-mini-2025-04-16 with data sharing | $100.00 / hour | $1.00 | $0.25 | $4.00 |

All models

Tokens used for model grading in reinforcement fine-tuning are billed at that model's per-token rate. Inference discounts are available if you enable data sharing when creating the fine-tune job. Learn more.

Next Model selection

Ask AI

Docs agent

Loading docs agent...