OpenAI

ChatGPT Rate Card (Enterprise token-based pricing)

Learn how token rates work for ChatGPT models and features across Chat, ChatGPT Work and Codex

Updated: 4 hours ago

Note: This rate card applies only to new Enterprise customers whose agreement specifies usage-based billing in USD. Rates are shown in U.S. dollars (USD) and, unless another billing unit is listed, are charged per 1 million input, cached input, and output tokens. Refer to your agreement for applicable rates, discounts, and other commercial terms.

If you are on an existing Enterprise/Edu plan, new Edu plan, or on a ChatGPT Business plan, refer to the ChatGPT Rate Card and Codex Rate card for credit based pricing.

If you are unsure which rate card applies to your workspace, contact your OpenAI representative.

Overview

This article outlines the current rates for ChatGPT features based on the token-based, flexible pricing structure for new Enterprise plans. It covers pricing for all features associated with Chat, Work and Codex usage.

Customers on credit-based pricing plans should refer to ChatGPT Rate Card (Credit based Pricing). For questions about your plan, please reach out to your OpenAI representative.

What are tokens?

Tokens are small units of information that OpenAI models process and generate. Token-based pricing measures usage across input tokens, cached input tokens, and output tokens. Your total charge is calculated by applying the applicable rate for the model and feature used to each token category, then adding those amounts together. You can learn more about how tokens are calculated here.

Chat models

Prices in USD per 1M tokens.

ModelInputCached inputOutput
GPT-5.6 SolMedium, High, Extra High$5.00$0.50$30.00
GPT-5.6 Sol Pro$5.00$0.50$30.00
GPT-5.5 Instant, Medium, High, Extra High$5.00$0.50$30.00
GPT-5.5 Pro$30.00N/A$180.00
GPT-5.5-Rosalind$12.50$1.25$75.00
o3$2.00$0.50$8.00
o3-pro$20.00N/A$80.00
GPT-5.3$1.75$0.175$14.00
  • Prices are expressed per million input, cached input, and output tokens. These rates replace ChatGPT message rates which were previously based on model and message count. Messages sent in chat will be based on token usage and model, so that usage pricing is standardized across all product surfaces.

  • The total cost of a request is calculated as follows: cost = (input tokens × input rate) + (cached-input tokens × cached-input rate) + (output tokens × output rate)

  • Instant uses GPT-5.5 Instant. Medium, High, and Extra High use different reasoning efforts, but selecting a higher reasoning effort does not increase the token price.

ChatGPT Work and Codex models

Prices in USD per 1M tokens.

ModelInputCached inputOutput
GPT-5.6 Sol$5.00$0.50$30.00
GPT-5.6 Terra$2.50$0.25$15.00
GPT-5.6 Luna$1.00$0.10$6.00
GPT-5.5$5.00$0.50$30.00
GPT-5.5 Cyber$12.50$1.25$75.00
GPT-5.5-Rosalind$12.50$1.25$75.00
GPT-5.4$2.50$0.25$15.00
GPT-5.4-Mini$0.75$0.075$4.52
GPT-5.3-Codex$1.75$0.175$14.00
GPT-5.2$1.75$0.175$14.00
GPT-5.3-Codex-SparkResearch previewResearch previewResearch preview

These rates apply to supported ChatGPT Work and Codex activity, including local tasks, cloud tasks, automations, code review, auto review, and delegated workers. Charges are based on the model used and the actual input, cached input, and output tokens consumed.

Image generation, Voice, web search, long context, fast mode, and regional processing may create separate or additional charges as described below.

Note:

  • Prices apply for all reasoning levels available for a given model (light, medium, high, extra high, ultra) unless otherwise noted.

  • The total cost of a request is calculated as follows: cost = (input tokens × input rate) + (cached-input tokens × cached-input rate) + (output tokens × output rate)

  • GPT-5.5 Cyber is available as a part of the OpenAI Daybreak/Trusted Access for Cyber program.

  • Code review uses GPT-5.3-Codex.

  • Auto review uses GPT-5.4.

  • GPT-5.3-Codex-Spark may be available in Codex as a research preview. Pricing for this model is not final.

Feature availability and applicability

The feature rates below apply wherever the listed feature is available across Chat, Work, and Codex, unless a section states otherwise. A single request may include both model-token charges and separate feature charges.

Core ChatGPT features

Deep research

Prices in USD per 1M tokens.

ModelInputCached inputOutput
o3-deep-research$10.00$2.50$40.00
o4-mini-deep-research$2.00$0.50$8.00

Images

These image-generation rates apply wherever GPT-Image-2.0 is available across ChatGPT, Work, and Codex.

Prices in USD per 1M tokens.

ModelInputCached inputOutput
GPT-Image-2.0 (image)$8.00$2.00$30.00
GPT-Image-2.0 (text)$5.00$1.25$10.00

Image generation usage is measured in tokens, with total cost based on input text tokens, input image tokens when editing an image, and image output tokens. Image output-token usage varies with the requested dimensions and quality, while edits that include reference images may require additional input tokens. Larger or higher-quality images generally use more tokens, although token usage can vary by resolution. For image generation cost estimates, use the calculator in the image generation guide.

Voice

FeatureRate
Voice in Chat$0.12 / minute+ Back end model (BEM) billed at token rates. BEM is GPT-5.5.
Voice in Work and Codex$0.12 / minute+ Back end model (BEM) and delegated workers billed at token rates. BEM is GPT-5.6 Terra.

Connected minutes and model activity are metered separately. A Voice session may therefore incur both the connected-minute charge and token charges for the backend model and any delegated workers.

When Voice starts a Codex task, the task is charged using the applicable ChatGPT Work and Codex model rates above. See ChatGPT Voice in Desktop pricing and limits for current availability, plan allowances, and usage limits.

Prices in USD per 1M tokens.

ModelInputCached inputOutput
gpt-4o-realtime-previewAudio$40.00$2.50$80.00
Text$5.00$2.50$20.00
gpt-4o-mini-realtime-previewAudio$10.00$0.30$20.00
Text$0.60$0.30$2.40

Realtime and audio model usage is generally priced per 1 million tokens, with separate rates for text and audio inputs, cached inputs, and outputs. Total cost is calculated by applying the relevant rate to each token category used during the interaction and adding the charges together. Some specialized models are priced using another unit, such as minutes of audio or characters generated, where noted.

Web search

FeatureRate
Web Search (all models)$10.00 / 1K web runs+ Search content tokens billed at model rates.
Image Web Search (all models)$10.00 / 1K web runs+ Search content tokens billed at model rates.

ChatGPT for Excel, PowerPoint, and Workspace Agents

Note: For customers on the token-based Enterprise plan, pricing is changing from a fixed number of credits per task to usage-based pricing calculated from tokens. ChatGPT for PowerPoint usage remains free for Business and Enterprise customers through August 6, 2026—after that date, charging starts under the same token-based pricing model as ChatGPT for Excel/Sheets.

ChatGPT for Excel / Sheets

Prices in USD per 1M tokens.

ModelInputCached inputOutput
GPT-5.6 Sol$5.00$0.50$30.00

ChatGPT for PowerPoint

Prices in USD per 1M tokens.

ModelInputCached inputOutput
GPT-5.6 Sol$5.00$0.50$30.00

ChatGPT Workspace Agents

Prices in USD per 1M tokens.

ModelInputCached inputOutput
GPT-5.6 Sol$5.00$0.50$30.00

Usage across Codex, ChatGPT Work, ChatGPT for Excel, ChatGPT for PowerPoint, and Workspace Agents is charged using the applicable model and feature rates in this article.

Note that these products share the same agentic usage when they are available on your plan.

Additional fees and multipliers

The following fees and multipliers may apply in addition to the standard model rate.

FeatureRate
Long context>272K input tokensInput: 2x Standard rateCached input: 2x Standard rateOutput: 1.5x Standard rate
Fast modeCodex and Work modes only; see Speed for more details.GPT-5.5: 2.5x Standard rateGPT-5.4: 2x Standard rate
Regional processing (data residency)See Your data guide for supported regions and processing details.1.1x Standard rate

Note:

  • Long context is only supported in GPT-5.6, GPT-5.5, and GPT-5.4 (Work and Codex tabs). Long context is not available in the Chat tab.

Monitoring usage and managing costs

You can monitor your Codex usage in the global admin console. For information about rate limits and reducing token consumption, see Codex pricing and usage limits and best practices for managing token consumption.

Actual costs vary based on the model used, task size, input and output mix, automations, fast mode, and the number of concurrent Codex instances.

Was this article helpful?