About Tiered Billing for GPT-5.6 Sol Ultra-Long Contexts
-
<p>According to OpenAI's latest billing rules, models such as GPT-5.6 Sol use tiered pricing for ultra-long context requests.</p><p>When the context of a single request exceeds 272K tokens, input, cached reads, and output may be billed at higher tier prices. This rule comes from OpenAI's official pricing and is not a temporary surcharge or abnormal charge from AI-ROUTER.</p><p></p><p>Using GPT-5.6 Sol's standard prices as an example:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Item Within 272K Over 272K</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ━━━━━━━━━━ ━━━━━━━━━━━━━━━━━━━ ━━━━━━━━━━━━━━━━━</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Input $5 / 1M tokens $10 / 1M tokens</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ────────── ─────────────────── ─────────────────</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Cached reads $0.50 / 1M tokens $1 / 1M tokens</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ────────── ─────────────────── ─────────────────</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Output $30 / 1M tokens $45 / 1M tokens</code></pre><p></p><p>The final cost is still calculated based on the specific group, account multiplier, and other billing settings.</p><p></p><p>If an x2 flag appears in usage records, it means that the request triggered tiered billing for long contexts. This flag mainly indicates that input or cached reads entered a higher tier; it does not mean that all charges for the request were simply doubled. Output may use a different tier multiplier.</p><p></p><p><strong> ## How to Avoid Triggering Tiered Billing</strong></p><p>If you do not need a 1M-token context window, we recommend adjusting Codex's context limit back to within 272K.</p><p></p><p>Open:</p><p> ~/.codex/config.toml</p><p></p><p>Change the 1M configuration from the original article:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model = "gpt-5.6-sol"</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_context_window = 1000000</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_auto_compact_token_limit = 900000</code></pre><p> </p><p>To:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model = "gpt-5.6-sol"</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_context_window = 272000</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_auto_compact_token_limit = 250000</code></pre><p> </p><p>Save the changes, restart Codex, and start a new session.</p><p></p><p>If you only want this to apply temporarily to a single session, you can use:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code>codex -m gpt-5.6-sol -c model_context_window=272000 -c model_auto_compact_token_limit=250000</code></pre><p> We also recommend:</p><p> - Enable automatic compaction promptly;</p><p> - Split ultra-long tasks across multiple sessions;</p><p> - Reduce the amount of tool output returned at once;</p><p> - Periodically summarize and clear older conversation content.</p><p></p><p>If you do need a 1M-token context window, you can continue using the original configuration, but usage beyond 272K will be billed according to OpenAI's higher tier pricing.</p><p></p><p>For the original configuration instructions, see: <a target="_blank" rel="noopener noreferrer nofollow" href="https://ai-router.dev/blog/post-g-677aa3e043912838-how-to-enable-a-1m-token-context-window-in-codex-for-gpt-5-6-sol">How to Enable a 1M-Token Context Window in Codex for GPT-5.6 Sol </a></p><p></p><p>Thank you for your understanding and support.</p>
Hello! It looks like you're interested in this conversation, but you don't have an account yet.
Getting fed up of having to scroll through the same posts each visit? When you register for an account, you'll always come back to exactly where you were before, and choose to be notified of new replies (either via email, or push notification). You'll also be able to save bookmarks and upvote posts to show your appreciation to other community members.
With your input, this post could be even better 💗
Register Login