메뉴
HN
Hacker News • 32일 전

OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)

IMP
5/10
핵심 요약

원문 보기
원문 보기 (영어)
Flagship models Our latest models Prices per 1M tokens. Standard Batch Flex Fast mode Standard Short context Long context Model Input Cached input Cache writes Output Input Cached input Cache writes Output gpt-5.6-sol $4.00 $0.40 $5.00 $20.00 $8.00 $0.80 $10.00 $30.00 gpt-5.6-terra $2.00 $0.20 $2.50 $12.00 $4.00 $0.40 $5.00 $18.00 gpt-5.6-luna $0.20 $0.02 $0.25 $1.20 $0.40 $0.04 $0.50 $1.80 Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details. OpenAI models in Amazon Bedrock are billed through AWS and may differ from direct OpenAI pricing. Priority processing was renamed Fast mode on July 30, 2026. You can use either service_tier: "priority" or service_tier: "fast" in your API requests. Learn more about Fast mode . GPT-5.6 Sol’s promotional pricing is available at least through November 21, 2026. All models Batch Short context Long context Model Input Cached input Cache writes Output Input Cached input Cache writes Output gpt-5.6-sol $2.00 $0.20 $2.50 $10.00 $4.00 $0.40 $5.00 $15.00 gpt-5.6-terra $1.00 $0.10 $1.25 $6.00 $2.00 $0.20 $2.50 $9.00 gpt-5.6-luna $0.10 $0.01 $0.125 $0.60 $0.20 $0.02 $0.25 $0.90 Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details. All models Flex Short context Long context Model Input Cached input Cache writes Output Input Cached input Cache writes Output gpt-5.6-sol $2.00 $0.20 $2.50 $10.00 $4.00 $0.40 $5.00 $15.00 gpt-5.6-terra $1.00 $0.10 $1.25 $6.00 $2.00 $0.20 $2.50 $9.00 gpt-5.6-luna $0.10 $0.01 $0.125 $0.60 $0.20 $0.02 $0.25 $0.90 Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details. All models Fast mode Short context Long context Model Input Cached input Cache writes Output Input Cached input Cache writes Output gpt-5.6-sol $8.00 $0.80 $10.00 $40.00 $16.00 $1.60 $20.00 $60.00 gpt-5.6-terra $4.00 $0.40 $5.00 $24.00 $8.00 $0.80 $10.00 $36.00 gpt-5.6-luna $0.40 $0.04 $0.50 $2.40 $0.80 $0.08 $1.00 $3.60 Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details. All models Cyber models Our latest Daybreak models. Prices per 1M tokens. Short context Long context Model Input Cached input Cache writes Output Input Cached input Cache writes Output gpt-5.6-sol $4.00 $0.40 $5.00 $20.00 $8.00 $0.80 $10.00 $30.00 gpt-5.6-cyber $12.50 $1.25 $15.625 $75.00 - - - - All models daybreak-blue-latest and daybreak-red-latest are aliases that currently point to gpt-5.6-sol and gpt-5.6-cyber , respectively. As new frontier models are released through the Daybreak program, these aliases will be updated to point to the latest models, with pricing adjusted to match each underlying model. Multimodal models Realtime and audio generation models Prices per 1M tokens unless noted. Model Modality Input Cached input Output / cost gpt-realtime-2.1 Audio $32.00 $0.40 $64.00 Text $4.00 $0.40 $24.00 Image $5.00 $0.50 - gpt-realtime-2.1-mini Audio $10.00 $0.30 $20.00 Text $0.60 $0.06 $2.40 Image $0.80 $0.08 - All models Image generation models Prices per 1M tokens. Standard Batch Standard For image generation cost estimates, use the calculator in the image generation guide. Model Modality Input Cached input Output gpt-image-2 Image $8.00 $2.00 $30.00 Text $5.00 $1.25 - All models Batch For image generation cost estimates, use the calculator in the image generation guide. Model Modality Input Cached input Output gpt-image-2 Image $4.00 $1.00 $15.00 Text $2.50 $0.625 - All models Video generation models Prices per second. Standard Batch Standard Model Size Portrait Landscape Price per second sora-2 720p 720x1280 1280x720 $0.10 sora-2-pro 720p 720x1280 1280x720 $0.30 1024p 1024x1792 1792x1024 $0.50 1080p 1080x1920 1920x1080 $0.70 Batch Model Size Portrait Landscape Price per second sora-2 720p 720x1280 1280x720 $0.05 sora-2-pro 720p 720x1280 1280x720 $0.15 1024p 1024x1792 1792x1024 $0.25 1080p 1080x1920 1920x1080 $0.35 Transcription models Prices per 1M tokens unless noted. Model Use case Input Output Estimated cost gpt-realtime-translate Live translation - - $0.034 / minute gpt-live-transcribe Live transcription - - $0.017 / minute gpt-realtime-whisper Live transcription - - $0.017 / minute gpt-transcribe Transcription - - $0.0045 / minute gpt-4o-transcribe Transcription $2.50 $10.00 $0.006 / minute gpt-4o-mini-transcribe Transcription $1.25 $5.00 $0.003 / minute All models Tools Tool Details Pricing Web search Web search (all models) $10.00 / 1k calls + Search content tokens billed at model rates. Image Web search (all models) $10.00 / 1k calls + Search content tokens billed at model rates. Web search preview (reasoning models, including gpt-5 , o-series ) $10.00 / 1k calls + Search content tokens billed at model rates. Web search preview (non-reasoning models) $25.00 / 1k calls + Search content tokens are free. Containers Hosted Shell and Code Interpreter 1 GB $0.03, 4 GB $0.12, 16 GB $0.48, 64 GB $1.92 per 20-minute session per container. File search Storage $0.10 / GB per day (1 GB free) Tool call $2.50 / 1k calls Agent Kit ChatKit file and image upload storage $0.10 / GB-day after 1 GB free per account per month Tokens used for built-in tools are billed at the chosen model's per-token rates. GB refers to binary gigabytes (also known as gibibytes), where 1 GB is 2^30 bytes. Web search content tokens are tokens retrieved from the search index and fed to the model alongside your prompt to generate an answer. For gpt-4o-mini and gpt-4.1-mini with the non-preview web search tool, search content tokens are billed as a fixed block of 8,000 input tokens per call. File search tool call pricing applies to the Responses API only. Container pricing includes Hosted Shell and Code Interpreter . Eligible container sessions will be billed by the minute, with a 5-minute minimum per session. Responses API, Chat Completions API, Realtime API, Batch API, and Assistants API are not priced separately. Tokens are billed at the chosen model's input and output rates. Specialized models Prices per 1M tokens. Standard Fast mode Standard Category Model Input Cached input Output ChatGPT chat-latest $5.00 $0.50 $30.00 Codex gpt-5.3-codex $1.75 $0.175 $14.00 All models Fast mode Category Model Input Cached input Output Codex gpt-5.3-codex $3.50 $0.35 $28.00 Finetuning Prices per 1M tokens. OpenAI is winding down the fine-tuning platform. The platform is no longer accessible to new users, but existing users of the fine-tuning platform will be able to create training jobs for the coming months. All fine-tuned models will remain available for inference until their base models are deprecated. The full timeline is here . Standard Batch Standard Model Training Input Cached input Output o4-mini-2025-04-16 $100.00 / hour $4.00 $1.00 $16.00 o4-mini-2025-04-16 with data sharing $100.00 / hour $2.00 $0.50 $8.00 All models Batch Model Training Input Cached input Output o4-mini-2025-04-16 $100.00 / hour $2.00 $0.50 $8.00 o4-mini-2025-04-16 with data sharing $100.00 / hour $1.00 $0.25 $4.00 All models Tokens used for model grading in reinforcement fine-tuning are billed at that model's per-token rate. Inference discounts are available if you enable data sharing when creating the fine-tune job. Learn more .
관련 소식