Home
API
Codex
ChatGPT
Resources
Start searching
API Dashboard
Get started
Overview
Quickstart
Models
Pricing
SDKs and CLI
Latest: GPT-5.5
Prompt guidance
Core concepts
Text generation
Code generation
Images and vision
Audio and speech
Structured output
Function calling
Responses API
Using tools
Agents SDK
Overview
Quickstart
Agent definitions
Models and providers
Running agents
Sandbox agents
Orchestration
Guardrails
Results and state
Integrations and observability
Evaluate agent workflows
Voice agents
ChatKit
Tools
Web search
MCP and Connectors
Skills
Shell
Computer use
File search and retrieval
Tool search
More tools
Run and scale
Conversation state
Background mode
Streaming
WebSocket mode
Webhooks
File inputs
Context management
Prompting
Reasoning
Evaluation
Red teaming
Realtime and audio
Overview
Voice agents
Live translation
Transcription
Speech generation
Realtime prompting guide
Connection methods
Realtime sessions
Specialized models
Image generation
Video generation
Deep research
Embeddings
Moderation
Going live
Production best practices
Workload identity federation
Deployment checklist
Amazon Bedrock
Latency optimization
Cost optimization
Accuracy optimization
Safety
Legacy APIs
Agent Builder
Evals
Fine-tuning
Assistants API
Resources
Terms and policies
Changelog
Your data
Permissions
Rate limits
IP egress ranges
Admin APIs
Deprecations
MCP for deep research
Developer mode
ChatGPT Actions
Flagship models
Multimodal models
Tools
Specialized models
Finetuning
Copy Page
Pricing
Copy Page

Flagship models

Our latest models
Prices per 1M tokens.
Standard
Batch
Flex
Priority
	
Short context
	
Long context

Model	Input	Cached input	
Output
	Input	Cached input	
Output


gpt-5.5
	
$5.00
	
$0.50
	
$30.00
	
$10.00
	
$1.00
	
$45.00


gpt-5.5-pro
	
$30.00
	
-
	
$180.00
	
$60.00
	
-
	
$270.00


gpt-5.4
	
$2.50
	
$0.25
	
$15.00
	
$5.00
	
$0.50
	
$22.50


gpt-5.4-mini
	
$0.75
	
$0.075
	
$4.50
	
-
	
-
	
-


gpt-5.4-nano
	
$0.20
	
$0.02
	
$1.25
	
-
	
-
	
-


gpt-5.4-pro
	
$30.00
	
-
	
$180.00
	
$60.00
	
-
	
$270.00
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details. OpenAI models in Amazon Bedrock are billed through AWS and may differ from direct OpenAI pricing.
All models

Multimodal models

Realtime and audio generation models

Prices per 1M tokens unless noted.

Model	Modality	Input	Cached input	Output / cost
gpt-realtime-2	Audio	$32.00	$0.40	$64.00
Text	$4.00	$0.40	$24.00
Image	$5.00	$0.50	-
gpt-realtime-translate	Audio	-	-	$0.034 / minute
gpt-realtime-whisper	Audio	-	-	$0.017 / minute
All models

Image generation models

Prices per 1M tokens.
Standard
Batch
For image generation cost estimates, use the calculator in the image generation guide.
Model	
Modality
	Input	Cached input	Output
gpt-image-2	Image	$8.00	$2.00	$30.00
Text	$5.00	$1.25	-
gpt-image-1.5	Image	$8.00	$2.00	$32.00
Text	$5.00	$1.25	$10.00
gpt-image-1-mini	Image	$2.50	$0.25	$8.00
Text	$2.00	$0.20	-
All models

Video generation models

Prices per second.
Standard
Batch
Model	Size	Portrait	Landscape	Price per second
sora-2	720p	720x1280	1280x720	$0.10
sora-2-pro	720p	720x1280	1280x720	$0.30
1024p	1024x1792	1792x1024	$0.50
1080p	1080x1920	1920x1080	$0.70

Transcription models

Prices per 1M tokens unless noted.

Model	Use case	Input	Output	Estimated cost
gpt-4o-transcribe	Transcription	$2.50	$10.00	$0.006 / minute
gpt-4o-mini-transcribe	Transcription	$1.25	$5.00	$0.003 / minute
All models

Tools

Tool	Details	Pricing
Web search	Web search (all models)	$10.00 / 1k calls
+ Search content tokens billed at model rates.
Image Web search (all models)	$10.00 / 1k calls
+ Search content tokens billed at model rates.
Web search preview (reasoning models, including gpt-5, o-series)	$10.00 / 1k calls
+ Search content tokens billed at model rates.
Web search preview (non-reasoning models)	$25.00 / 1k calls
+ Search content tokens are free.
Containers	Hosted Shell and Code Interpreter	1 GB $0.03, 4 GB $0.12, 16 GB $0.48, 64 GB $1.92 per 20-minute session per container.
File search	Storage	$0.10 / GB per day (1 GB free)
Tool call	$2.50 / 1k calls
Agent Kit	ChatKit file and image upload storage	$0.10 / GB-day after 1 GB free per account per month
Tokens used for built-in tools are billed at the chosen model's per-token rates. GB refers to binary gigabytes (also known as gibibytes), where 1 GB is 2^30 bytes. Web search content tokens are tokens retrieved from the search index and fed to the model alongside your prompt to generate an answer. For gpt-4o-mini and gpt-4.1-mini with the non-preview web search tool, search content tokens are billed as a fixed block of 8,000 input tokens per call. File search tool call pricing applies to the Responses API only. Container pricing includes Hosted Shell and Code Interpreter. Eligible container sessions will be billed by the minute, with a 5-minute minimum per session. Responses API, Chat Completions API, Realtime API, Batch API, and Assistants API are not priced separately. Tokens are billed at the chosen model's input and output rates.

Specialized models

Prices per 1M tokens.
Standard
Batch
Priority
Category	Model	Input	Cached input	Output
ChatGPT	chat-latest	$5.00	$0.50	$30.00
Codex	gpt-5.3-codex	$1.75	$0.175	$14.00
Cyber	gpt-5.4-cyber	-	-	-
All models

Finetuning

Prices per 1M tokens.

OpenAI is winding down the fine-tuning platform. The platform is no longer accessible to new users, but existing users of the fine-tuning platform will be able to create training jobs for the coming months.




All fine-tuned models will remain available for inference until their base models are deprecated. The full timeline is here.

Standard
Batch
Model	Training	Input	Cached input	Output
o4-mini-2025-04-16	$100.00 / hour	$4.00	$1.00	$16.00
o4-mini-2025-04-16
with data sharing	$100.00 / hour	$2.00	$0.50	$8.00
All models
Tokens used for model grading in reinforcement fine-tuning are billed at that model's per-token rate. Inference discounts are available if you enable data sharing when creating the fine-tune job. Learn more.
Ask AI
Docs agent

Loading docs agent...