GPT-6 Sol API: Pricing, the 272K Cliff and Setup
GPT-6 Sol API pricing is $2/$10 per 1M tokens at or below 272K. Above that the whole request bills $4/$15. Cache reads are $0.16. Live as gpt-6-sol.

GPT-6 Sol is OpenAI's coding and agent model in the GPT-6 family, live on RouterPlex as gpt-6-sol. GPT-6 Sol API pricing is $2 per 1M input tokens and $10 per 1M output tokens at or below 272,000 input tokens. Cross that line and the whole request bills $4 input and $15 output per 1M.
That is one fifth of GPT-6 Astra ($10 / $50) and twenty times GPT-6 Luna ($0.10 / $0.50) on the input side. OpenAI shipped both Sol and Luna on 23 September 2026. Astra stays the flagship.


Sources: OpenAI's Introducing GPT-6 Sol and Luna (23 September 2026), the GPT-6 Sol model card, and the live RouterPlex catalog, checked 23 September 2026. Scores below are OpenAI's reported results. Rates change. Confirm the GPT-6 Sol price page before you lock a budget.
GPT-6 Sol API pricing #
USD per 1M tokens. The first two rows are what RouterPlex bills. The cache row is the catalog cache-read rate, which is separate from OpenAI's published cached-input line.
| Tier | Input | Output | Cache read |
|---|---|---|---|
| Standard (≤272K input) | $2.00 | $10.00 | $0.16 |
| Long context (>272K input) | $4.00 | $15.00 | $0.32 |
OpenAI's model card, same day, lists a different cache sheet for the same sticker:
| OpenAI line | Per 1M |
|---|---|
| Input | $2.00 |
| Cached input | $0.20 |
| Cache writes | $2.50 |
| Output | $10.00 |
RouterPlex bills cache reads at $0.16 per 1M ($0.32 above 272K). That is the number on the live catalog. It is lower than OpenAI's published $0.20 cached-input rate. RouterPlex does not publish a separate cache-write price, a Batch price, or a Flex price. OpenAI's card prices Batch and Flex at 50% of Standard, Fast at 2×, and regional processing at a 10% uplift where that processing is available. This key bills the standard list rate in the table above.
A 40,000-in / 2,000-out turn with no cache costs $0.10. The same turn is $0.005 on gpt-6-luna, $0.50 on gpt-6-astra, and $0.26 on gpt-5.6-sol at the catalog rate ($5 / $30).
A $5 prepaid balance covers 2.5 million Sol input tokens, or 500,000 Sol output tokens, if a workload used only one side. Real requests mix both. Reasoning tokens bill as output.
The 272K cliff applies to the whole request #
Once input tokens go past 272,000, the higher rate applies to every token in that request. There is no blended tail.
271,000-token prompt on gpt-6-sol → 271,000 × $2.00/1M = $0.542273,000-token prompt on gpt-6-sol → 273,000 × $4.00/1M = $1.092
Two thousand extra input tokens about double the input bill, and output on that same request moves from $10 to $15 per 1M. If a coding agent is anywhere near 272K, trimming the prompt below the threshold is worth more than a small cache tweak. The same rule is on Astra ($20 / $75) and Luna ($0.20 / $0.75).
GPT-6 Sol vs Luna, Astra, and GPT-5.6 Sol #
List prices on RouterPlex, USD per 1M tokens, and the cost of a 40,000-token prompt returning 2,000 tokens with no cache. The GPT-5.6 Sol row is the catalog rate, which is above OpenAI's $4 / $20 promotional list.
| Model | Input | Output | Cache read | 40K / 2K turn | Max input here |
|---|---|---|---|---|---|
gpt-6-luna | $0.10 | $0.50 | $0.014 | $0.005 | 922K |
gpt-6-sol | $2.00 | $10.00 | $0.16 | $0.100 | 922K |
gpt-5.6-sol | $5.00 | $30.00 | $0.40 | $0.260 | 258K |
gpt-6-astra | $10.00 | $50.00 | $0.80 | $0.500 | 922K |
OpenAI's 23 September 2026 table shows GPT-6 Sol at $2 / $10 against a GPT-5.6 Sol promotional price of $4 / $20, and labels that a 50% reduction. RouterPlex bills the new Sol list. It still bills GPT-5.6 Sol at $5 / $30. If you are costing a migration off the old Sol ID, use $0.10 versus $0.26 on this turn, and read the GPT-5.6 Sol vs Terra vs Luna page for the older family.
Sol is the model to benchmark when the task is coding, tool use, or a multi-step agent and Astra's $10 / $50 is more than the task can return. Luna is the model when the same workflow is clerical and you can accept the smaller model. The cheapest models ledger ranks the rest of the catalog by agent-turn cost. Kimi K3 sits nearby at $3 / $15 if you want a non-OpenAI comparison at a similar sticker.
What OpenAI measured #
These figures are from OpenAI's 23 September 2026 announcement. They are vendor-run, at the effort level OpenAI names, against competitor scores OpenAI took from public reports. RouterPlex has not re-run them.
| Eval | GPT-6 Sol | Comparison OpenAI cites | Cost note from OpenAI |
|---|---|---|---|
| AutomationBench 1.0.6 (Zapier, 47 tools) | 33.2% at xhigh, $0.27/task | Opus 5 max 26.9%; Astra low 30.3%; Fable 5.1 with Opus 5 fallback max 31.4% | Opus 5 was 11.1× Sol's cost per task. Astra low was 3.9×. The Fable 5.1 cost omits fallbacks on about 40% of tasks |
| Agents' Last Exam | 56.4% at max | Above Opus 5's highest score in that eval | 60% lower cost per task than that Opus 5 run |
| DeepSWE v1.1 | 68.8% at max | Fable 5 highest 69.9% at xhigh | About 80% lower cost per task |
| OSWorld 2.0 offline (v2026.08.08, partial reward) | 60.5% at xhigh | Opus 5 medium 60.3% | About 80% lower cost per task |
| FrontierCode | Improves on GPT-5.6 Sol | Matches Fable 5.1 xhigh | OpenAI did not publish the exact scores in the post |
| Internal factuality | About half as many mistakes as its predecessor | Approaches Astra | Error-inducing de-identified chats, not typical usage, and not length-controlled |
Luna's own scores, including DeepSWE 66.6% at max effort, are on the GPT-6 Luna pricing guide. Alignment notes, including coding-deception rates on deliberately hard cases, are in the GPT-6 Astra system card. Those evals are challenging cases, and OpenAI says they are not typical failure rates.
What the model card documents #
| Capability | GPT-6 Sol |
|---|---|
| Model ID | gpt-6-sol |
| Context window | 1,050,000 tokens |
| Max input | 922,000 tokens |
| Max output | 128,000 tokens |
| Knowledge cutoff | 20 April 2026 |
| Input / output | Text and image in, text out |
| Reasoning effort | none, low, medium (default), high, xhigh, max |
| On RouterPlex | Reasoning, vision, tool calling |
OpenAI's card says the Responses API is the path for built-in tools, and that Chat Completions supports function calling only with reasoning_effort set to none. On RouterPlex, call Sol through Chat Completions at https://api.routerplex.com/v1. Set that effort when the request includes tools. Hosted OpenAI tools on the card (web search, file search, code interpreter, computer use, and the rest) are an OpenAI Responses surface. Do not assume they are attached to this gateway route.
OpenAI also publishes standard rate limits on the card: tier 1 is 500 RPM and 500K TPM, and tier 5 is 15,000 RPM and 40M TPM. Those are OpenAI's limits. RouterPlex does not publish a separate per-model RPM table for this ID. The cap you set is the hard budget on the key.
Call GPT-6 Sol on RouterPlex #
One OpenAI-compatible key, prepaid, with a hard spend limit. The balance stops at $0. There is no per-token markup and no top-up fee.
curl https://api.routerplex.com/v1/chat/completions \-H "Authorization: Bearer $ROUTERPLEX_API_KEY" \-H "Content-Type: application/json" \-d '{"model": "gpt-6-sol","messages": [{"role": "user", "content": "Name the riskiest assumption in this migration plan."}]}'
import osfrom openai import OpenAIclient = OpenAI(api_key=os.environ["ROUTERPLEX_API_KEY"],base_url="https://api.routerplex.com/v1",)response = client.chat.completions.create(model="gpt-6-sol",messages=[{"role": "user", "content": "Design a rollback-safe deployment plan."}],)print(response.choices[0].message.content)
Claude Code and the Anthropic SDK use ANTHROPIC_BASE_URL=https://api.routerplex.com and model gpt-6-sol. The Claude Code setup guide has the full configuration. Cursor can override the OpenAI base URL to https://api.routerplex.com/v1. Codex CLI can use a custom provider.
Live prices: GPT-6 Sol, GPT-6 Luna, GPT-6 Astra, and the OpenAI hub.
When Sol is the right ID #
- Coding agents and hard professional tasks where you want the GPT-6 tier and Astra's $0.50 turn is more than the task returns. Budget $2 / $10 and stay under 272K unless you mean to pay the cliff.
- A migration off GPT-5.6 Sol. The new ID is cheaper on this catalog ($0.10 versus $0.26 on the sample turn) and the input window is 922K rather than 258K. Re-run your own task set before you delete the old ID.
- High-volume extraction, classification, and triage. Use GPT-6 Luna at $0.10 / $0.50. Sol is twenty times the input price.
- The newest flagship. Use GPT-6 Astra. OpenAI still calls Astra the best model in the family.
Start a $5 RouterPlex test, put gpt-6-sol on a budgeted key, and send one live request. The same key reaches Luna, Astra, and the rest of the catalog. Gateways that fund credits with a percentage fee are compared on RouterPlex vs OpenRouter.
Common questions
Frequently asked questions
How much does the GPT-6 Sol API cost?
OpenAI lists GPT-6 Sol at $2 per 1M input tokens and $10 per 1M output tokens for prompts at or below 272,000 input tokens. Prompts above 272,000 bill $4 input and $15 output per 1M for the whole request. RouterPlex bills those list rates on model ID gpt-6-sol with no per-token markup.
What is the GPT-6 Sol context window?
OpenAI documents a 1,050,000-token context window, 922,000 maximum input tokens, and 128,000 maximum output tokens. The RouterPlex catalog lists 922,000 input tokens and a 128,000 output cap. Knowledge cutoff is 20 April 2026.
What happens to GPT-6 Sol pricing above 272K tokens?
The higher rate applies to every token in that request. A 271,000-token prompt is about $0.542 of input. A 273,000-token prompt is about $1.092 of input, because the whole request moves from $2 to $4 per 1M, and output on that request moves from $10 to $15 per 1M.
How much does RouterPlex charge for GPT-6 Sol cached input?
The live catalog bills cache reads at $0.16 per 1M tokens, and $0.32 per 1M once the prompt crosses 272K. OpenAI's own sheet lists cached input at $0.20 and cache writes at $2.50. RouterPlex publishes the cache-read rate only.
How does GPT-6 Sol compare with GPT-6 Luna and Astra on price?
On a 40,000-in / 2,000-out turn with no cache, Sol is $0.10, Luna is $0.005, and Astra is $0.50. Sol is the coding and agent tier. Luna is the high-volume tier. Astra remains the flagship at $10 / $50.
Is GPT-6 Sol cheaper than GPT-5.6 Sol?
Yes, on this catalog. GPT-6 Sol is $2 / $10. RouterPlex bills GPT-5.6 Sol at $5 / $30. OpenAI's GPT-5.6 Sol promotional list is $4 / $20 through at least 21 November 2026, and this route does not bill that promo. The same 40K / 2K turn is $0.10 on GPT-6 Sol and $0.26 on GPT-5.6 Sol here.
Can I use GPT-6 Sol with the OpenAI SDK?
Yes. Point the OpenAI SDK at https://api.routerplex.com/v1 with model ID gpt-6-sol. Claude Code uses ANTHROPIC_BASE_URL=https://api.routerplex.com and the same model ID. OpenAI's model card says Chat Completions function calling on this model works only when reasoning_effort is none. Use that setting if you call tools through Chat Completions.
Run the smallest paid test.
Add $5, cap the key, and verify the result with your own workload. No subscription, and credit never expires — a first top-up of $25+ is matched with $25 extra.



