$5.00 input · $30.00 output per 1M tokens — OpenAI's flagship ceiling for the hardest requests.
GPT-5.6 Sol is OpenAI's flagship: $5.00 input and $30.00 output per million tokens, with cached input at $0.50. In 2026 the value winners in nearly every price band are mid-tier or budget models — Sol exists for the rare request that justifies paying for the ceiling.
GPT-5.6 Sol · USD official list price
| Tier | Rate |
|---|---|
| Input | $5.00 |
| Output | $30.00 |
| Cache hit | $0.50 |
DeepSeek lists official prices in CNY; USD uses ≈ 6.9 exchange rate
Standard-tier list price for input and output per million tokens.
Cached input is 10× cheaper than uncached, rewarding stable prompt prefixes.
Offline workloads qualify for OpenAI's Batch API discount, halving output cost.
The hard 10% of requests: rare, complex reasoning and high-stakes generation where a mid-tier or budget model demonstrably fails. The other 90% should never pay flagship prices.
Published API pricing as of September 2026, standard tier, USD per million tokens. GPT-5.6 models all qualify for the 50% Batch API discount for offline workloads.
Against DeepSeek V4 Pro (¥9.00/$27.00 peak), Sol's $30.00 output is roughly 10% pricier than V4 Pro's peak output in USD terms yet represents the same flagship tier — with V4 Pro winning on off-peak economics and 1M context.
GPT-5.6 Sol bills $5.00 per million input tokens and $30.00 per million output tokens on the standard tier. Cached input is $0.50 per million. OpenAI also offers a 50% Batch API discount for offline workloads, bringing output to $15.00.
Only for the hard 10% of requests: rare, complex reasoning where a mid-tier or budget model demonstrably fails. In 2026 the value winners in nearly every price band are mid-tier or budget models — Sol exists for the rare request that justifies it.
Yes. GPT-5.6 Sol charges $0.50 per million tokens for cached input (10× cheaper than the $5.00 uncached input rate), so keeping a stable prompt prefix materially lowers effective input cost.
Compare this model against the rest of the lineup.
Model your own workload — input/output tokens, cache hit rate, and peak-hour share — against any model mix with the interactive pricing calculator.
Open the Pricing CalculatorWe are using these tools ourselves for Development / Deployment. Check out for more details.
You will receive $5 when subscribed, which directly offsets the 1st month's Go fee.
Could be the cheapest option for calling DeepSeek V4 Flash model in the world.
RReceive $300 to test out VVultr platform
Referred user must be active 30+ days and spend $10–$25
Deploy anything without the complexity
Connect your repo, Railway handles the rest.
Cloud servers, cloud databases, COS, CDN, SMS and other cloud products are on special offer now.
AWS high-quality alternative, we use COS as a replacement for S3
¥16 universal voucher for all platforms, valid for 180 days from the date of receipt
A comprehensive product matrix supporting the full-process implementation of AI applications.
One of the best AI terminal tools in the world
Interestingly, we can directly use various top-tier models in Warp without leaving the terminal.