For the complete documentation index, see llms.txt. Markdown versions of documentation pages are available by appending
.md to the page URL.
ChatGPT
Home
API
Codex
Docs
Guides, concepts, and product docs for Codex
Use cases
Example workflows and tasks teams can take on with ChatGPT or Codex
Docs
Use cases
Resources
ChatGPT
Plugins
Extend ChatGPT and Codex
Workspace Agents
Trigger published ChatGPT workspace agents
Commerce
Build commerce flows in ChatGPT
Ads
Publish and measure ads in ChatGPT
Resources
Showcase
Demo apps to get inspired
Blog
Learnings and experiences from developers
Cookbook
Notebook examples for building with OpenAI models
Learn
Docs, videos, and demo apps for building with OpenAI
Community
Programs, meetups, and support for builders
Start searching
API Dashboard
Try ChatGPT
Overview Models Agents Tools Voice & Audio Production API reference
Search the API docs
Search docs
Suggested
responses createreasoning_effortrealtimeprompt caching
Primary navigation
API Codex ChatGPT Docs Use cases Resources Resources
Search docs
Suggested
responses createreasoning_effortrealtimeprompt caching
Overview Models Agents Tools Voice & Audio Production API reference
OverviewModelsAgentsToolsVoice & AudioProductionAPI referenceDocs sectionModels
Home
Get started
Quickstart
Using GPT-5.6
Key concepts
Core concepts
Responses API
Conversation state
Background mode
Streaming
WebSocket mode
Multi-agent
Webhooks
File inputs
Compaction
Counting tokens
SDKs and CLI
OpenAI SDK
OpenAI CLI
Resources
Changelog
Deprecations
Supported countries
OpenAI Crawlers
Terms and policies
Legacy APIs
Agent Builder
Overview
Migration guide
Node reference
Safety in building agents
Evals
Getting started
Working with evals
Prompt optimizer
External models
Best practices
Graders
Fine-tuning
Optimization cycle
Supervised fine-tuning
Vision fine-tuning
Direct preference optimization
Reinforcement fine-tuning
RFT use cases
Best practices
Assistants API
Migration guide
Model catalog
Choose a model
Pricing
Model selection
Text and code
Text generation
Code generation
Structured output
Prompting
Overview
Prompt engineering
Citation formatting
Migration guide
Prompt generation
Frontend prompting
Reasoning
Reasoning models
Reasoning best practices
Images and video
Images and vision
Image input cost calculator
Image generation
Video generation
Realtime and audio
Audio and speech
Overview
Voice agents
Specialized models
Deep research
Embeddings
Moderation
Overview
Agents SDK
Quickstart
Agent definitions
Models and providers
Running agents
Sandbox agents
Orchestration
Guardrails
Results and state
Integrations and observability
Evaluate agent workflows
ChatKit
Overview
Customize
Widgets
Actions
Advanced integrations
Overview
Function calling
Search and retrieval
Web search
File search
Retrieval
Connect tools and data
MCP and Connectors
Secure MCP Tunnel
Build tool workflows
Skills
Tool search
Programmatic tool calling
Computer and code
Shell
Computer use
Apply Patch
Local shell
Code interpreter
Media
Image generation
Overview
Get started
Voice agents
Live translation
Realtime prompting guide
Audio
Audio and speech
Transcription
File transcription
Realtime transcription
Speech generation
Connection methods
WebRTC
WebSocket
SIP
Sessions and operations
Managing conversations
Voice activity detection
Realtime with tools
Webhooks and server-side controls
Managing costs
Go live
Production best practices
Deployment checklist
Performance and quality
Latency optimization
Predicted Outputs
Fast mode
Accuracy optimization
Cost and throughput
Cost optimization
Prompt caching
Batch
Flex processing
Safety and governance
Safety best practices
Red teaming
Safety checks
Cybersecurity checks
Under 18 API Guidance
CSAM guidance
Content provenance
Your data
Permissions
Infrastructure and access
Terraform provider
Overview
Projects and access
Service accounts
Rate limits and spend
Model, tool, and data controls
Import and reconciliation
Private Link
IP allowlist
Mutual TLS
Workload identity federation
Codex setup
Federation rules
Admin API
X.509 certificates
Kubernetes
AWS
Microsoft Azure
Google Cloud
Oracle Cloud Infrastructure
GitHub Actions
SPIFFE
IP egress ranges
Amazon Bedrock
Operations
Rate limits
Spend limits
Admin APIs
Error codes
Docs Use cases
DocsUse casesDocs sectionDocs
Plugins Workspace Agents Commerce Ads
PluginsWorkspace AgentsCommerceAdsDocs sectionSelect...
Home
Quickstart
Core concepts
Plugin architecture
Skills
MCP server
Plan
Brainstorm use cases
Define tools
Build
Build an MCP server
Add UI to your MCP server (optional)
Authenticate users
Build skills
Package your plugin
Examples
Test and publish
Connect and test your plugin
Submit and publish
Submission error reference
Conversion specs
Restaurant reservation spec
Get Quote spec
Product checkout spec
Guides
UI guidelines
Optimize Metadata
Submit a Claude Code plugin
Security & Privacy
Troubleshooting
Resources
Changelog
Plugin guidelines
MCP server review requirements
Plugin UI reference
Checkout API reference
Home
Get started
Trigger workspace agent runs
Authenticate with Workspace Agent access tokens
Home
Guides
Get started
Best practices
File Upload
Overview
Products
API
Overview
Feeds
Products
Promotions
Ads Overview
Measurement
Measurement Pixel
Multiple Pixels (Advanced)
Image Tag
Conversions API
Supported Events
Advertiser API
Overview
API Partner Setup
Quickstart
Bulk API
Product Feeds
Delta Feeds API
Campaign Targeting
Conversion-Optimized Campaigns
Custom Audiences
API Reference
Authentication
Ad Account
Campaigns
Ad Groups
Ads
Insights
Files
Conversion Setup
Overview Features Configuration Developers Security Administration Use Cases Resources
OverviewFeaturesConfigurationDevelopersSecurityAdministrationUse CasesResourcesDocs sectionOverview
Home
Get started
Quickstart
Use ChatGPT
Get started with Work
Import from another agent
Foundations
Prompting
Personalize ChatGPT
Skills & Plugins
Permissions
Explore
What's new
Models
Pricing
Glossary
Available on
ChatGPT desktop app
Remote
ChatGPT on the web
Codex CLI
Codex IDE extension
Codex cloud
Releases
Changelog
Feature Maturity
Open Source
Overview
Workflows
Projects and chats
Sites
Visualizations
Scheduled tasks
Long-running work
Notifications
Pets
Codex Micro
Capabilities
Browser
Computer use
Voice
Plugins
Web search
Image generation
Image inputs
Appshots
Browser extension
Work with files
Reference
Commands
Slash commands
Settings
Troubleshooting
Overview
Customization
Overview
Memories
Computer History
Config file
Config Basics
Advanced Config
Config Reference
Environment Variables
Sample Config
Agent configuration
AGENTS.md
Subagents
Speed
Rules
Extend ChatGPT and Codex
Record & Replay
MCP
Linux
Desktop app
Windows
Desktop app
Windows sandbox
WSL
Overview
Development workflows
Code review
Integrated terminal
Extend and automate
Build skills
Build plugins
Site tools (WebMCP)
Hooks
Environments
Modes
Local environments
Cloud environment
Git worktrees
Build with Codex
Codex SDK
App Server
MCP Server
GitHub Action
Non-interactive mode
Third-party integrations
GitHub
GitLab (Beta)
Slack
Linear
Reference
CLI customization
Developer commands
Developer settings
Overview
Permissions
Profiles
Sandboxing
Auto-review
Agent approvals & security
Internet access
Codex Security
Overview
Codex Security plugin
Quickstart
Run a security scan
Run a deep scan
Review code changes
Use the Security workbench
Triage a backlog
Fix findings
Propose security hardening
Write vulnerability reports
Export and track findings
Changelog
Codex Security CLI
Quickstart
Run bulk scans
Run scans in CI
GitLab CI/CD
Reference
FAQ
TypeScript SDK
Codex Security cloud
Setup
Security Review
Improving the threat model
FAQ
Cyber safety
Models & Trusted Access
Recommended configuration
Overview
Getting started
Admin rollout guide
ChatGPT Work Overview
ChatGPT Work cloud security
ChatGPT Work admin FAQ
Identity and authentication
Authentication overview
Workload identity
Personal Access Tokens
Service accounts
Workspace access, policy, and models
Groups and provisioning
Roles and workspace permissions
GPTs and Sharing
Managed configuration
Prisma AIRS
HIPAA configuration
Workspace model availability
Plugin and connector controls
Plugin controls
Skill controls
Usage, governance, and compliance
Governance
Workspace analytics
Analytics API
Compliance API and audit events
Deployment and model providers
Manage app updates
Windows app deployment
Remote connections
Amazon Bedrock
Explore use cases
Collections
Home
Videos
Showcase
OpenAI Academy
Online trainings
Community
Codex Ambassadors
Codex for Students
Codex for Open Source
Meetups
Blog
Company blog
Developer blog
Explore use cases
Collections
Home
Videos
Showcase
OpenAI Academy
Online trainings
Community
Codex Ambassadors
Codex for Students
Codex for Open Source
Meetups
Blog
Company blog
Developer blog
Showcase Blog Cookbook Learn Community
ShowcaseBlogCookbookLearnCommunityDocs sectionSelect...
All posts
Recent
Automating repetitive work at OpenAI with Codex
Meet the winners of OpenAI Build Week
Scaling cyber defenders with Daybreak
Codex as a platform: build on the open agent harness
Custom Code Review rules for Codex
Topics
General
API
Apps SDK
Audio
Codex
Home
Topics
Agents
Evals
Multimodal
Text
Guardrails
Optimization
ChatGPT
Codex
gpt-oss
Contribute
Cookbook on GitHub
Home
OpenAI Developers plugin
Docs MCP
Categories
Demo apps
Videos
Topics
Agents
Audio & Voice
Computer Use
Codex
Evals
gpt-oss
Fine-tuning
Image generation
Scaling
Tools
Video generation
Community
Programs
Codex Ambassadors
Codex for Students
Codex for Open Source
OpenAI for Startups
Events
Meetups
Spaces
Developer Forum
Discord
Reddit
X
API Dashboard
Try ChatGPT
Model catalog
Choose a model
Pricing
Model selection
Text and code
Text generation
Code generation
Structured output
Prompting
Overview
Prompt engineering
Citation formatting
Migration guide
Prompt generation
Frontend prompting
Reasoning
Reasoning models
Reasoning best practices
Images and video
Images and vision
Image input cost calculator
Image generation
Video generation
Realtime and audio
Audio and speech
Overview
Voice agents
Specialized models
Deep research
Embeddings
Moderation
Copy Page
Pricing
Copy Page
Flagship models
Our latest models
Prices per 1M tokens.
StandardBatchFlexFast mode
Standard
Short context
Long context
Model
Input
Cached input
Cache writes
Output
Input
Cached input
Cache writes
Output
gpt-5.6-sol
$4.00
$0.40
$5.00
$20.00
$8.00
$0.80
$10.00
$30.00
gpt-5.6-terra
$2.00
$0.20
$2.50
$12.00
$4.00
$0.40
$5.00
$18.00
gpt-5.6-luna
$0.20
$0.02
$0.25
$1.20
$0.40
$0.04
$0.50
$1.80
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details. OpenAI models in Amazon Bedrock are billed through AWS and may differ from direct OpenAI pricing.
Priority processing was renamed Fast mode on July 30, 2026. You can use either service_tier: "priority" or service_tier: "fast" in your API requests. Learn more about Fast mode.
GPT-5.6 Sol’s promotional pricing is available at least through November 21, 2026.
All models
Batch
Short context
Long context
Model
Input
Cached input
Cache writes
Output
Input
Cached input
Cache writes
Output
gpt-5.6-sol
$2.00
$0.20
$2.50
$10.00
$4.00
$0.40
$5.00
$15.00
gpt-5.6-terra
$1.00
$0.10
$1.25
$6.00
$2.00
$0.20
$2.50
$9.00
gpt-5.6-luna
$0.10
$0.01
$0.125
$0.60
$0.20
$0.02
$0.25
$0.90
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.
All models
Flex
Short context
Long context
Model
Input
Cached input
Cache writes
Output
Input
Cached input
Cache writes
Output
gpt-5.6-sol
$2.00
$0.20
$2.50
$10.00
$4.00
$0.40
$5.00
$15.00
gpt-5.6-terra
$1.00
$0.10
$1.25
$6.00
$2.00
$0.20
$2.50
$9.00
gpt-5.6-luna
$0.10
$0.01
$0.125
$0.60
$0.20
$0.02
$0.25
$0.90
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.
All models
Fast mode
Short context
Long context
Model
Input
Cached input
Cache writes
Output
Input
Cached input
Cache writes
Output
gpt-5.6-sol
$8.00
$0.80
$10.00
$40.00
$16.00
$1.60
$20.00
$60.00
gpt-5.6-terra
$4.00
$0.40
$5.00
$24.00
$8.00
$0.80
$10.00
$36.00
gpt-5.6-luna
$0.40
$0.04
$0.50
$2.40
$0.80
$0.08
$1.00
$3.60
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.
All models
Cyber models
Our latest Daybreak models.
Prices per 1M tokens.
Short context
Long context
Model
Input
Cached input
Cache writes
Output
Input
Cached input
Cache writes
Output
gpt-5.6-sol
$4.00
$0.40
$5.00
$20.00
$8.00
$0.80
$10.00
$30.00
gpt-5.6-cyber
$12.50
$1.25
$15.625
$75.00
-
-
-
-
All models
daybreak-blue-latest and daybreak-red-latest are
aliases that currently point to gpt-5.6-sol and
gpt-5.6-cyber, respectively. As new frontier models are released
through the Daybreak program, these aliases will be updated to point to the
latest models, with pricing adjusted to match each underlying model.
Multimodal models
To estimate vision model input costs, use the image input cost
calculator.
Realtime and audio generation models
Prices per 1M tokens unless noted.
Model
Modality
Input
Cached input
Output / cost
gpt-realtime-2.1
Audio
$32.00
$0.40
$64.00
Text
$4.00
$0.40
$24.00
Image
$5.00
$0.50
-
gpt-realtime-2.1-mini
Audio
$10.00
$0.30
$20.00
Text
$0.60
$0.06
$2.40
Image
$0.80
$0.08
-
All models
Image generation models
Prices per 1M tokens.
StandardBatch
Standard
For image generation cost estimates, use the calculator in the image generation guide.
Model
Modality
Input
Cached input
Output
gpt-image-2
Image
$8.00
$2.00
$30.00
Text
$5.00
$1.25
-
All models
Batch
For image generation cost estimates, use the calculator in the image generation guide.
Model
Modality
Input
Cached input
Output
gpt-image-2
Image
$4.00
$1.00
$15.00
Text
$2.50
$0.625
-
All models
Video generation models
Prices per second.
StandardBatch
Standard
Model
Size
Portrait
Landscape
Price per second
sora-2
720p
720x1280
1280x720
$0.10
sora-2-pro
720p
720x1280
1280x720
$0.30
1024p
1024x1792
1792x1024
$0.50
1080p
1080x1920
1920x1080
$0.70
Batch
Model
Size
Portrait
Landscape
Price per second
sora-2
720p
720x1280
1280x720
$0.05
sora-2-pro
720p
720x1280
1280x720
$0.15
1024p
1024x1792
1792x1024
$0.25
1080p
1080x1920
1920x1080
$0.35
Transcription models
Prices per 1M tokens unless noted.
Model
Use case
Input
Output
Estimated cost
gpt-realtime-translate
Live translation
-
-
$0.034 / minute
gpt-live-transcribe
Live transcription
-
-
$0.017 / minute
gpt-realtime-whisper
Live transcription
-
-
$0.017 / minute
gpt-transcribe
Transcription
-
-
$0.0045 / minute
gpt-4o-transcribe
Transcription
$2.50
$10.00
$0.006 / minute
gpt-4o-mini-transcribe
Transcription
$1.25
$5.00
$0.003 / minute
All models
Tools
Tool
Details
Pricing
Web search
Web search (all models)
$10.00 / 1k calls
+ Search content tokens billed at model rates.
Image Web search (all models)
$10.00 / 1k calls
+ Search content tokens billed at model rates.
Web search preview (reasoning models, including gpt-5, o-series)
$10.00 / 1k calls
+ Search content tokens billed at model rates.
Web search preview (non-reasoning models)
$25.00 / 1k calls
+ Search content tokens are free.
Containers
Hosted Shell and Code Interpreter
1 GB $0.03, 4 GB $0.12, 16 GB $0.48, 64 GB $1.92 per 20-minute session per container.
File search
Storage
$0.10 / GB per day (1 GB free)
Tool call
$2.50 / 1k calls
Agent Kit
ChatKit file and image upload storage
$0.10 / GB-day after 1 GB free per account per month
Tokens used for built-in tools are billed at the chosen model's per-token rates. GB refers to binary gigabytes (also known as gibibytes), where 1 GB is 2^30 bytes. Web search content tokens are tokens retrieved from the search index and fed to the model alongside your prompt to generate an answer. For gpt-4o-mini and gpt-4.1-mini with the non-preview web search tool, search content tokens are billed as a fixed block of 8,000 input tokens per call. File search tool call pricing applies to the Responses API only. Container pricing includes Hosted Shell and Code Interpreter. Eligible container sessions will be billed by the minute, with a 5-minute minimum per session. Responses API, Chat Completions API, Realtime API, Batch API, and Assistants API are not priced separately. Tokens are billed at the chosen model's input and output rates.
Specialized models
Prices per 1M tokens.
StandardFast mode
Standard
Category
Model
Input
Cached input
Output
ChatGPT
chat-latest
$5.00
$0.50
$30.00
Codex
gpt-5.3-codex
$1.75
$0.175
$14.00
All models
Fast mode
Category
Model
Input
Cached input
Output
Codex
gpt-5.3-codex
$3.50
$0.35
$28.00
Finetuning
Prices per 1M tokens.
OpenAI is winding down the fine-tuning platform. The platform is no longer
accessible to new users, but existing users of the fine-tuning platform
will be able to create training jobs for the coming months.
All fine-tuned models will remain available for inference until their base
models are deprecated. The full timeline is
here.
StandardBatch
Standard
Model
Training
Input
Cached input
Output
o4-mini-2025-04-16
$100.00 / hour
$4.00
$1.00
$16.00
o4-mini-2025-04-16
with data sharing
$100.00 / hour
$2.00
$0.50
$8.00
All models
Batch
Model
Training
Input
Cached input
Output
o4-mini-2025-04-16
$100.00 / hour
$2.00
$0.50
$8.00
o4-mini-2025-04-16
with data sharing
$100.00 / hour
$1.00
$0.25
$4.00
All models
Tokens used for model grading in reinforcement fine-tuning are billed at that model's per-token rate. Inference discounts are available if you enable data sharing when creating the fine-tune job. Learn more.
Next
Model selection
Ask AI
Docs agent
Loading docs agent...