GPT-5.6 Solの超長コンテキスト段階料金について
-
<p>OpenAIの最新の料金ルールに基づき、GPT-5.6 Solなどのモデルでは、超長コンテキストリクエストに段階料金が適用されます。</p><p>1回のリクエストでコンテキストが272Kトークンを超えると、入力、キャッシュ読み取り、出力に対して、より高い段階の料金が適用される場合があります。このルールはOpenAI公式の料金体系によるものであり、AI-ROUTERによる一時的な上乗せ料金や異常な請求ではありません。</p><p></p><p>GPT-5.6 Solの標準料金を例にすると、次のようになります。</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> 項目 272K以内 272K超過</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ━━━━━━━━━━ ━━━━━━━━━━━━━━━━━━━ ━━━━━━━━━━━━━━━━━</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> 入力 $5 / 1Mトークン $10 / 1Mトークン</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ────────── ─────────────────── ─────────────────</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> キャッシュ読み取り $0.50 / 1Mトークン $1 / 1Mトークン</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ────────── ─────────────────── ─────────────────</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> 出力 $30 / 1Mトークン $45 / 1Mトークン</code></pre><p></p><p>最終的な料金は、具体的なグループ、アカウント倍率、その他の料金設定に基づいて計算されます。</p><p></p><p>利用記録にx2フラグが表示されている場合、そのリクエストで長いコンテキスト向けの段階料金が適用されたことを示します。このフラグは主に、入力やキャッシュ読み取りがより高い段階に移行したことを示すものであり、リクエスト全体のすべての料金が単純に2倍になったことを意味するわけではありません。出力には異なる段階倍率が適用される場合があります。</p><p></p><p><strong> ## 段階料金の適用を避ける方法</strong></p><p>1Mトークンのコンテキストウィンドウが不要な場合は、Codexのコンテキスト上限を272K以内に戻すことをおすすめします。</p><p></p><p>次のファイルを開きます。</p><p> ~/.codex/config.toml</p><p></p><p>記事内で以前使用していた1M設定:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model = "gpt-5.6-sol"</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_context_window = 1000000</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_auto_compact_token_limit = 900000</code></pre><p> </p><p>次のように変更します。</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model = "gpt-5.6-sol"</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_context_window = 272000</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_auto_compact_token_limit = 250000</code></pre><p> </p><p>保存後、Codexを再起動して新しいセッションを開始します。</p><p></p><p>単一のセッションに一時的にのみ適用したい場合は、次のコマンドを使用できます。</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code>codex -m gpt-5.6-sol -c model_context_window=272000 -c model_auto_compact_token_limit=250000</code></pre><p> あわせて、次の対策もおすすめします。</p><p> - 自動圧縮を適時有効にする;</p><p> - 長時間のタスクを複数のセッションに分割する;</p><p> - 1回の応答で大量のツール出力を返さないようにする;</p><p> - 定期的に要約し、古い会話内容を整理する。</p><p></p><p>1Mトークンのコンテキストウィンドウが本当に必要な場合は、元の設定を引き続き使用できます。ただし、272Kを超えた部分には、OpenAIの高い段階の料金が適用されます。</p><p></p><p>元の設定については、次の記事を参照してください:<a target="_blank" rel="noopener noreferrer nofollow" href="https://ai-router.dev/blog/post-g-677aa3e043912838-how-to-enable-a-1m-token-context-window-in-codex-for-gpt-5-6-sol">CodexでGPT-5.6 Solの100万トークンコンテキストウィンドウを有効にする方法 </a></p><p></p><p>ご理解とご支援に感謝いたします。</p>
Hello! It looks like you're interested in this conversation, but you don't have an account yet.
Getting fed up of having to scroll through the same posts each visit? When you register for an account, you'll always come back to exactly where you were before, and choose to be notified of new replies (either via email, or push notification). You'll also be able to save bookmarks and upvote posts to show your appreciation to other community members.
With your input, this post could be even better 💗
Register Login