About Tiered Billing for Extra-Long Contexts in GPT-5.6 Sol
-
<p>According to OpenAI's latest billing rules, models such as GPT-5.6 Sol use tiered pricing for extra-long context requests.</p><p>When the context of a single request exceeds 272K tokens, input, cached reads, and output may be billed at higher-tier prices. This rule comes from OpenAI's official pricing and does not represent a temporary surcharge or abnormal charge from AI-ROUTER.</p><p></p><p>Using GPT-5.6 Sol's standard pricing as an example:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Item Within 272K Over 272K</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ━━━━━━━━━━ ━━━━━━━━━━━━━━━━━━━ ━━━━━━━━━━━━━━━━━</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Input $5 / 1M tokens $10 / 1M tokens</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ────────── ─────────────────── ─────────────────</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Cached read $0.50 / 1M tokens $1 / 1M tokens</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ────────── ─────────────────── ─────────────────</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Output $30 / 1M tokens $45 / 1M tokens</code></pre><p></p><p>The final cost is still calculated based on the specific group, account multiplier, and other billing settings.</p><p></p><p>An x2 indicator in usage records means that the request triggered long-context tiered billing. This indicator mainly means that input and cached reads entered a higher tier; it does not mean that all costs for the entire request are simply doubled. Output may use a different tier multiplier.</p><p></p><p><strong> ## How to Avoid Triggering Tiered Billing</strong></p><p>If you do not need a 1M-token context window, it is recommended that you adjust Codex's context limit back to 272K or below.</p><p></p><p>Open:</p><p> ~/.codex/config.toml</p><p></p><p>Change the 1M configuration from the original article:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model = "gpt-5.6-sol"</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_context_window = 1000000</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_auto_compact_token_limit = 900000</code></pre><p> </p><p>to:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model = "gpt-5.6-sol"</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_context_window = 272000</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_auto_compact_token_limit = 250000</code></pre><p> </p><p>Save the file, restart Codex, and start a new session.</p><p></p><p>If you only want this to apply temporarily to a single session, you can use:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code>codex -m gpt-5.6-sol -c model_context_window=272000 -c model_auto_compact_token_limit=250000</code></pre><p> It is also recommended that you:</p><p> - Enable automatic compaction promptly;</p><p> - Split extra-long tasks across multiple sessions;</p><p> - Reduce the amount of tool output returned at once;</p><p> - Regularly summarize and clear older conversation content.</p><p></p><p>If you do need a 1M-token context window, you can continue using the original configuration, but usage beyond 272K will be billed according to OpenAI's higher-tier pricing.</p><p></p><p>For the original configuration instructions, see: <a target="_blank" rel="noopener noreferrer nofollow" href="https://ai-router.dev/blog/post-g-677aa3e043912838-how-to-enable-a-1m-token-context-window-in-codex-for-gpt-5-6-sol">How to Enable a 1M-Token Context Window in Codex for GPT-5.6 Sol </a></p><p></p><p>Thank you for your understanding and support.</p>
Hello! It looks like you're interested in this conversation, but you don't have an account yet.
Getting fed up of having to scroll through the same posts each visit? When you register for an account, you'll always come back to exactly where you were before, and choose to be notified of new replies (either via email, or push notification). You'll also be able to save bookmarks and upvote posts to show your appreciation to other community members.
With your input, this post could be even better 💗
Register Login