OpenAI API DocsCommunity translation · Official structure

Pricing

Pricing

Pricing information for the OpenAI platform.

English source
This English page is rendered from the official Markdown mirror in this repository.View on OpenAI ↗

For the complete documentation index, see llms.txt. Markdown versions of documentation pages are available by appending .md to the page URL.

Flagship models

Our latest models

Prices per 1M tokens.

Standard

Standard pricing data

ModelShort context inputShort context cached inputShort context cache writesShort context outputLong context inputLong context cached inputLong context cache writesLong context output
gpt-5.6-sol$4.00$0.40$5.00$20.00$8.00$0.80$10.00$30.00
gpt-5.6-terra$2.00$0.20$2.50$12.00$4.00$0.40$5.00$18.00
gpt-5.6-luna$0.20$0.02$0.25$1.20$0.40$0.04$0.50$1.80
gpt-5.5 (<272K context length)$5.00$0.50-$30.00$10.00$1.00-$45.00
gpt-5.5-pro (<272K context length)$30.00--$180.00$60.00--$270.00
gpt-5.4 (<272K context length)$2.50$0.25-$15.00$5.00$0.50-$22.50
gpt-5.4-mini$0.75$0.075-$4.50----
gpt-5.4-nano$0.20$0.02-$1.25----
gpt-5.4-pro (<272K context length)$30.00--$180.00$60.00--$270.00
gpt-5.2$1.75$0.175-$14.00----
gpt-5.2-pro$21.00--$168.00----
gpt-5.1$1.25$0.125-$10.00----
gpt-5$1.25$0.125-$10.00----
gpt-5-mini$0.25$0.025-$2.00----
gpt-5-nano$0.05$0.005-$0.40----
gpt-5-pro$15.00--$120.00----
gpt-4.1$2.00$0.50-$8.00----
gpt-4.1-mini$0.40$0.10-$1.60----
gpt-4.1-nano$0.10$0.025-$0.40----
gpt-4o$2.50$1.25-$10.00----
gpt-4o-2024-05-13$5.00--$15.00----
gpt-4o-mini$0.15$0.075-$0.60----
o1$15.00$7.50-$60.00----
o1-pro$150.00--$600.00----
o3-pro$20.00--$80.00----
o3$2.00$0.50-$8.00----
o4-mini$1.10$0.275-$4.40----
o3-mini$1.10$0.55-$4.40----
gpt-4-turbo-2024-04-09$10.00--$30.00----
gpt-4-0613$30.00--$60.00----
gpt-3.5-turbo$0.50--$1.50----
gpt-3.5-turbo-0125$0.50--$1.50----
gpt-3.5-turbo-1106$1.00--$2.00----
gpt-3.5-turbo-instruct$1.50--$2.00----
davinci-002$2.00--$2.00----
babbage-002$0.40--$0.40----

Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details. OpenAI models in Amazon Bedrock are billed through AWS and may differ from direct OpenAI pricing. Priority processing was renamed Fast mode on July 30, 2026. You can use either service_tier: "priority" or service_tier: "fast" in your API requests. Learn more about Fast mode. GPT-5.6 Sol’s promotional pricing is available at least through November 21, 2026.

Batch

Batch pricing data

ModelShort context inputShort context cached inputShort context cache writesShort context outputLong context inputLong context cached inputLong context cache writesLong context output
gpt-5.6-sol$2.00$0.20$2.50$10.00$4.00$0.40$5.00$15.00
gpt-5.6-terra$1.00$0.10$1.25$6.00$2.00$0.20$2.50$9.00
gpt-5.6-luna$0.10$0.01$0.125$0.60$0.20$0.02$0.25$0.90
gpt-5.5 (<272K context length)$2.50$0.25-$15.00$5.00$0.50-$22.50
gpt-5.5-pro (<272K context length)$15.00--$90.00----
gpt-5.4 (<272K context length)$1.25$0.13-$7.50$2.50$0.25-$11.25
gpt-5.4-mini$0.375$0.0375-$2.25----
gpt-5.4-nano$0.10$0.01-$0.625----
gpt-5.4-pro (<272K context length)$15.00--$90.00$30.00--$135.00
gpt-5.2$0.875$0.0875-$7.00----
gpt-5.2-pro$10.50--$84.00----
gpt-5.1$0.625$0.0625-$5.00----
gpt-5$0.625$0.0625-$5.00----
gpt-5-mini$0.125$0.0125-$1.00----
gpt-5-nano$0.025$0.0025-$0.20----
gpt-5-pro$7.50--$60.00----
gpt-4.1$1.00--$4.00----
gpt-4.1-mini$0.20--$0.80----
gpt-4.1-nano$0.05--$0.20----
gpt-4o$1.25--$5.00----
gpt-4o-2024-05-13$2.50--$7.50----
gpt-4o-mini$0.075--$0.30----
o1$7.50--$30.00----
o1-pro$75.00--$300.00----
o3-pro$10.00--$40.00----
o3$1.00--$4.00----
o4-mini$0.55--$2.20----
o3-mini$0.55--$2.20----
gpt-4-turbo-2024-04-09$5.00--$15.00----
gpt-4-0613$15.00--$30.00----
gpt-3.5-turbo-0125$0.25--$0.75----
gpt-3.5-turbo-1106$1.00--$2.00----
davinci-002$1.00--$1.00----
babbage-002$0.20--$0.20----

Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.

Flex

Flex pricing data

ModelShort context inputShort context cached inputShort context cache writesShort context outputLong context inputLong context cached inputLong context cache writesLong context output
gpt-5.6-sol$2.00$0.20$2.50$10.00$4.00$0.40$5.00$15.00
gpt-5.6-terra$1.00$0.10$1.25$6.00$2.00$0.20$2.50$9.00
gpt-5.6-luna$0.10$0.01$0.125$0.60$0.20$0.02$0.25$0.90
gpt-5.5 (<272K context length)$2.50$0.25-$15.00$5.00$0.50-$22.50
gpt-5.5-pro (<272K context length)$15.00--$90.00----
gpt-5.4 (<272K context length)$1.25$0.13-$7.50$2.50$0.25-$11.25
gpt-5.4-mini$0.375$0.0375-$2.25----
gpt-5.4-nano$0.10$0.01-$0.625----
gpt-5.4-pro (<272K context length)$15.00--$90.00$30.00--$135.00
gpt-5.2$0.875$0.0875-$7.00----
gpt-5.1$0.625$0.0625-$5.00----
gpt-5$0.625$0.0625-$5.00----
gpt-5-mini$0.125$0.0125-$1.00----
gpt-5-nano$0.025$0.0025-$0.20----
o3$1.00$0.25-$4.00----
o4-mini$0.55$0.138-$2.20----

Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.

Fast mode

Fast pricing data

ModelShort context inputShort context cached inputShort context cache writesShort context outputLong context inputLong context cached inputLong context cache writesLong context output
gpt-5.6-sol$8.00$0.80$10.00$40.00$16.00$1.60$20.00$60.00
gpt-5.6-terra$4.00$0.40$5.00$24.00$8.00$0.80$10.00$36.00
gpt-5.6-luna$0.40$0.04$0.50$2.40$0.80$0.08$1.00$3.60
gpt-5.5 (<272K context length)$12.50$1.25-$75.00----
gpt-5.4 (<272K context length)$5.00$0.50-$30.00----
gpt-5.4-mini$1.50$0.15-$9.00----
gpt-5.2$3.50$0.35-$28.00----
gpt-5.1$2.50$0.25-$20.00----
gpt-5$2.50$0.25-$20.00----
gpt-5-mini$0.45$0.045-$3.60----
gpt-4.1$3.50$0.875-$14.00----
gpt-4.1-mini$0.70$0.175-$2.80----
gpt-4.1-nano$0.20$0.05-$0.80----
gpt-4o$4.25$2.125-$17.00----
gpt-4o-2024-05-13$8.75--$26.25----
gpt-4o-mini$0.25$0.125-$1.00----
o3$3.50$0.875-$14.00----
o4-mini$2.00$0.50-$8.00----

Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.

Cyber models

Our latest Daybreak models.

Prices per 1M tokens.

Grouped Pricing Table data

ModelShort context inputShort context cached inputShort context cache writesShort context outputLong context inputLong context cached inputLong context cache writesLong context output
gpt-5.6-sol$4.00$0.40$5.00$20.00$8.00$0.80$10.00$30.00
gpt-5.6-cyber$12.50$1.25$15.625$75.00----
gpt-5.5-cyber$12.50$1.25-$75.00----
gpt-5.4-cyber--------

daybreak-blue-latest and daybreak-red-latest are aliases that currently point to gpt-5.6-sol and gpt-5.6-cyber, respectively. As new frontier models are released through the Daybreak program, these aliases will be updated to point to the latest models, with pricing adjusted to match each underlying model.

Multimodal models

Realtime and audio generation models

Prices per 1M tokens unless noted.

Grouped Pricing Table data

ModelModalityInputCached inputOutput / cost
gpt-realtime-2.1Audio$32.00$0.40$64.00
gpt-realtime-2.1Text$4.00$0.40$24.00
gpt-realtime-2.1Image$5.00$0.50-
gpt-realtime-2.1-miniAudio$10.00$0.30$20.00
gpt-realtime-2.1-miniText$0.60$0.06$2.40
gpt-realtime-2.1-miniImage$0.80$0.08-
gpt-realtime-2Audio$32.00$0.40$64.00
gpt-realtime-2Text$4.00$0.40$24.00
gpt-realtime-2Image$5.00$0.50-
gpt-realtime-1.5Audio$32.00$0.40$64.00
gpt-realtime-1.5Text$4.00$0.40$16.00
gpt-realtime-1.5Image$5.00$0.50-
gpt-realtime-miniAudio$10.00$0.30$20.00
gpt-realtime-miniText$0.60$0.06$2.40
gpt-realtime-miniImage$0.80$0.08-
gpt-realtimeAudio$32.00$0.40$64.00
gpt-realtimeText$4.00$0.40$16.00
gpt-realtimeImage$5.00$0.50-
gpt-audio-1.5Audio$32.00-$64.00
gpt-audio-1.5Text$2.50-$10.00
gpt-audio-miniAudio$10.00-$20.00
gpt-audio-miniText$0.60-$2.40
gpt-audioAudio$32.00-$64.00
gpt-audioText$2.50-$10.00
gpt-4o-mini-ttsAudio--$12.00
gpt-4o-mini-ttsText$0.60--
tts-1Text$15.00 / 1M characters--
tts-1-hdText$30.00 / 1M characters--

Image generation models

Prices per 1M tokens.

Standard

  For image generation cost estimates, use the calculator in the image generation guide.
  

Grouped Pricing Table data

ModelModalityInputCached inputOutput
gpt-image-2Image$8.00$2.00$30.00
gpt-image-2Text$5.00$1.25-
gpt-image-1.5Image$8.00$2.00$32.00
gpt-image-1.5Text$5.00$1.25$10.00
gpt-image-1-miniImage$2.50$0.25$8.00
gpt-image-1-miniText$2.00$0.20-
gpt-image-1Image$10.00$2.50$40.00
gpt-image-1Text$5.00$1.25-
chatgpt-image-latestImage$8.00$2.00$32.00
chatgpt-image-latestText$5.00$1.25$10.00

Batch

  For image generation cost estimates, use the calculator in the image generation guide.
  

Grouped Pricing Table data

ModelModalityInputCached inputOutput
gpt-image-2Image$4.00$1.00$15.00
gpt-image-2Text$2.50$0.625-
gpt-image-1.5Image$4.00$1.00$16.00
gpt-image-1.5Text$2.50$0.63$5.00
gpt-image-1-miniImage$1.25$0.13$4.00
gpt-image-1-miniText$1.00$0.10-
gpt-image-1Image$5.00$1.25$20.00
gpt-image-1Text$2.50$0.63-
chatgpt-image-latestImage$4.00$1.00$16.00
chatgpt-image-latestText$2.50$0.63$5.00

Video generation models

Prices per second.

Standard

Grouped Pricing Table data

ModelSizePortraitLandscapePrice per second
sora-2720p720x12801280x720$0.10
sora-2-pro720p720x12801280x720$0.30
sora-2-pro1024p1024x17921792x1024$0.50
sora-2-pro1080p1080x19201920x1080$0.70

Batch

Grouped Pricing Table data

ModelSizePortraitLandscapePrice per second
sora-2720p720x12801280x720$0.05
sora-2-pro720p720x12801280x720$0.15
sora-2-pro1024p1024x17921792x1024$0.25
sora-2-pro1080p1080x19201920x1080$0.35

Transcription models

Prices per 1M tokens unless noted.

Grouped Pricing Table data

ModelUse caseInputOutputEstimated cost
gpt-realtime-translateLive translation--$0.034 / minute
gpt-live-transcribeLive transcription--$0.017 / minute
gpt-realtime-whisperLive transcription--$0.017 / minute
gpt-transcribeTranscription--$0.0045 / minute
gpt-4o-transcribeTranscription$2.50$10.00$0.006 / minute
gpt-4o-mini-transcribeTranscription$1.25$5.00$0.003 / minute
gpt-4o-transcribe-diarizeTranscription + diarization$2.50$10.00$0.006 / minute
WhisperTranscription--$0.006 / minute

Tools

Grouped Pricing Table data

ToolDetailsPricing
Web searchWeb search (all models)$10.00 / 1k calls + Search content tokens billed at model rates.
Web searchImage Web search (all models)$10.00 / 1k calls + Search content tokens billed at model rates.
Web searchWeb search preview (reasoning models, including gpt-5, o-series)$10.00 / 1k calls + Search content tokens billed at model rates.
Web searchWeb search preview (non-reasoning models)$25.00 / 1k calls + Search content tokens are free.
ContainersHosted Shell and Code Interpreter1 GB $0.03, 4 GB $0.12, 16 GB $0.48, 64 GB $1.92 per 20-minute session per container.
File searchStorage$0.10 / GB per day (1 GB free)
File searchTool call$2.50 / 1k calls
Agent KitChatKit file and image upload storage$0.10 / GB-day after 1 GB free per account per month

$10.00 / 1k calls + Search content tokens billed at model rates.

Web search preview (reasoning models, including gpt-5, o-series)

$25.00 / 1k calls + Search content tokens are free.

Hosted Shell and Code Interpreter

Tokens used for built-in tools are billed at the chosen model's per-token rates. GB refers to binary gigabytes (also known as gibibytes), where 1 GB is 2^30 bytes. Web search content tokens are tokens retrieved from the search index and fed to the model alongside your prompt to generate an answer. For gpt-4o-mini and gpt-4.1-mini with the non-preview web search tool, search content tokens are billed as a fixed block of 8,000 input tokens per call. File search tool call pricing applies to the Responses API only. Container pricing includes Hosted Shell and Code Interpreter. Eligible container sessions will be billed by the minute, with a 5-minute minimum per session. Responses API, Chat Completions API, Realtime API, Batch API, and Assistants API are not priced separately. Tokens are billed at the chosen model's input and output rates.

Specialized models

Prices per 1M tokens.

Standard

Grouped Pricing Table data

CategoryModelInputCached inputOutput
ChatGPTchat-latest$5.00$0.50$30.00
Codexgpt-5.3-codex$1.75$0.175$14.00
Searchgpt-5-search-api$1.25$0.125$10.00
Embeddingtext-embedding-3-small$0.02--
Embeddingtext-embedding-3-large$0.13--
Embeddingtext-embedding-ada-002$0.10--
Moderationomni-moderation-latestFree--

Fast mode

Grouped Pricing Table data

CategoryModelInputCached inputOutput
Codexgpt-5.3-codex$3.50$0.35$28.00

Finetuning

Prices per 1M tokens.

OpenAI is winding down the fine-tuning platform. The platform is no longer
  accessible to new users, but existing users of the fine-tuning platform
  will be able to create training jobs for the coming months.
  

  All fine-tuned models will remain available for inference until their base
  models are deprecated. The full timeline is
  [here](https://developers.openai.com/api/docs/deprecations#update-to-openais-self-serve-fine-tuning).

Standard

Pricing Table data

ModelTrainingInputCached inputOutput
o4-mini-2025-04-16$100.00 / hour$4.00$1.00$16.00
o4-mini-2025-04-16 (data sharing)$100.00 / hour$2.00$0.50$8.00
gpt-4.1-2025-04-14$25.00$3.00$0.75$12.00
gpt-4.1-mini-2025-04-14$5.00$0.80$0.20$3.20
gpt-4.1-nano-2025-04-14$1.50$0.20$0.05$0.80
gpt-4o-2024-08-06$25.00$3.75$1.875$15.00
gpt-4o-mini-2024-07-18$3.00$0.30$0.15$1.20
gpt-3.5-turbo (legacy)$8.00$3.00-$6.00
davinci-002 (legacy)$6.00$12.00-$12.00
babbage-002 (legacy)$0.40$1.60-$1.60

Batch

Pricing Table data

ModelTrainingInputCached inputOutput
o4-mini-2025-04-16$100.00 / hour$2.00$0.50$8.00
o4-mini-2025-04-16 (data sharing)$100.00 / hour$1.00$0.25$4.00
gpt-4.1-2025-04-14$25.00$1.50$0.50$6.00
gpt-4.1-mini-2025-04-14$5.00$0.40$0.10$1.60
gpt-4.1-nano-2025-04-14$1.50$0.10$0.025$0.40
gpt-4o-2024-08-06$25.00$2.225$0.90$12.50
gpt-4o-mini-2024-07-18$3.00$0.15$0.075$0.60
gpt-3.5-turbo (legacy)$8.00$1.50-$3.00
davinci-002 (legacy)$6.00$6.00-$6.00
babbage-002 (legacy)$0.40$0.80-$0.90

Tokens used for model grading in reinforcement fine-tuning are billed at that model's per-token rate. Inference discounts are available if you enable data sharing when creating the fine-tune job. Learn more.