最初に: 「Sol / Luna」は2世代あります
2026年9月22日以降、「Sol」「Luna」という名前のモデルは2世代分存在します。検索結果や比較記事で混在しやすいので、まずここを押さえてください。
| 世代 | モデル名 | 料金(入力/出力) | 当サイトのページ |
|---|---|---|---|
| 旧(GPT-5.6世代) | GPT-5.6 Sol | $4 / $20 | /gpt-5-6-sol/ |
| 旧(GPT-5.6世代) | GPT-5.6 Luna | $0.20 / $1.20 | /gpt-5-6-luna/ |
| 新(GPT-6世代) | GPT-6 Sol | $2 / $10 | 本ページ |
| 新(GPT-6世代) | GPT-6 Luna | $0.10 / $0.50 | 本ページ |
OpenAIの表現は「reducing API prices for Sol and Luna by 50% compared with their GPT-5.6 promotional pricing」=同じ名前の前世代からの置き換えです。別モデルが併存しているのではなく、後継が旧価格を半額にしたという関係です。
料金(USD / 1Mトークン)
OpenAI公式発表の料金表と、OpenAI APIドキュメントのモデルページで確認。
| モデル | 入力 | 出力 | キャッシュ入力 | 前世代からの変化 |
|---|---|---|---|---|
| GPT-6 Sol | $2.00 | $10.00 | $0.20 | $4/$20 → 50%値下げ |
| GPT-6 Luna | $0.10 | $0.50 | $0.01 | $0.20/$1.20 → 50%値下げ |
※ キャッシュ入力はAPIドキュメントの記載値です(Sol $0.20、Luna $0.01)。公式発表の本文には「discounts of 90% on cached input-token reads」とあり、$2.00の10%=$0.20、$0.10の10%=$0.01 と整合します。
一次情報: OpenAI「Introducing GPT-6 Sol and Luna」(料金表)· developers.openai.com GPT-6 Sol · GPT-6 Luna(2026-09-25確認)
図: 世代ごとの単価
← 横にスクロールできます →
モデル仕様
提供状況(公式の逐語)
- 「available in ChatGPT Work and Codex starting today for all Plus, Pro, Business, Enterprise, and Edu users」
- 「Free and Go users can access GPT-6 Luna in the desktop app」
- 「These models are not yet available in Chat」(=通常のChatGPTチャットでは未提供)
- ChatGPTへの展開は当日中に段階的に実施。表示されない場合は時間をおいて再確認するよう案内されています。
出典: OpenAI APIドキュメント GPT-6 Sol · GPT-6 Luna · 公式発表のAvailability節
公式ベンチマーク6項目 — 運営元と条件つきで
OpenAI公式発表に掲載された数値です。OpenAIは各ベンチの出典を明示しており、第三者運営のベンチが半分以上を占めます。当サイトは「誰が運営しているか」を併記します。
| ベンチマーク | 運営元 | GPT-6 Sol の結果 | GPT-6 Luna の結果 |
|---|---|---|---|
| AutomationBench 1.0.6 | Zapier | (xhigh) 33.2% / $0.27 per task | (high) 前世代比 +5.4pt・タスク単価58%減 |
| Agents' Last Exam V1 | agents-last-exam.org | (max) 56.4%(Claude Opus 5 の最高スコア超え・コスト60%減) | — |
| FrontierCode 1.1 Main | Cognition | GPT-5.6 Sol から大幅改善、Fable 5.1 (xhigh) に並ぶ(はるかに低コスト) | — |
| DeepSWE 1.1 | Datacurve | (max) 68.8% — Fable 5 の最高69.9%と1.1pt差・コスト約80%減 | (max) 66.6% — Opus 5 / Fable 5 の medium相当 |
| OSWorld 2.0(offline・partial) | XLANG Lab | (xhigh) 60.5% vs Claude Opus 5 (medium) 60.3%・コスト約80%減 | (max) GPT-5.6 Sol (medium) 超え・コスト1/10 |
| Factuality(社内評価) | OpenAI(社内) | 前世代の約半分のミス | 高エフォート時、GPT-5.6 Sol と同等を約100分の1のコストで |
各ベンチの条件(公式注記の逐語要約)
| ベンチマーク | 条件・注記 |
|---|---|
| AutomationBench 1.0.6 | 「AI agents are tested on end-to-end workflows using 47 tools across sales, marketing, operations, support, finance, and HR」。比較表はSol (xhigh) 33.2% / $0.27、Astra (low) 30.3% / 3.9倍、Claude Opus 5 (max) 26.9% / 11.1倍、Fable 5.1 + Opus 5フォールバック (max) 31.4% / 8.9倍超。OpenAI自身の注記:「The datapoint for Claude Fable 5.1 understates its actual cost, as it omits the cost of the Opus 5 fallbacks, which occurred on ~40% of tasks」。 |
| Agents' Last Exam V1 | 「long-horizon, economically valuable tasks spanning 55 sub-industries」。企業の計算機上で行われる専門業務の大半をカバー。 |
| FrontierCode 1.1 Main | 「graded not only on correctness but also “mergeability”: e.g., test quality, scope discipline, code style, and adherence to codebase standards」=正しさだけでなくマージ可能性を採点。 |
| DeepSWE 1.1 | 「AI agents solve original, long-horizon software engineering tasks」=オリジナルの長時間タスク。 |
| OSWorld 2.0 | 「long-horizon computer-use workflows spanning everyday and professional tasks. We report the partial reward on the offline set from the v2026.08.08 release」=オフライン版の部分報酬。 |
| Factuality | 「de-identified ChatGPT conversations where users had flagged a factual error」を基にした社内評価。OpenAIの注記:「These error-inducing conversations are not representative of typical usage」「Scores are not controlled for length」。 |
出典: OpenAI「Introducing GPT-6 Sol and Luna」(各ベンチ節と注記・2026-09-25確認)
もうひとつの本題: プロンプトキャッシュの改善
値下げと同じくらい実務に効くのが、GPT-6向けのキャッシュ改善です。OpenAIは「キャッシュ入力トークンの読み出しで90%割引」に加え、既定でキャッシュヒット率が上がるよう改善したと説明しています。
| 機能 | 内容(公式発表より) |
|---|---|
| モニタリング | Prompt Caching Dashboardで入力のうちどれだけがキャッシュされたかを時系列で確認。 |
| 診断 | diagnostics toolがキャッシュの取り逃しを説明し、修正点を示す。 |
| キャッシュを壊さず設定変更 | 推論エフォートの変更やツールの有効化/無効化をしても、以前のコンテキストがキャッシュ再利用のために保持される。 |
| キャッシュ境界の指定 | 明示的なブレークポイントで、キャッシュするプレフィックスの終端を開発者が選べる。 |
| 実績(GitHub) | GitHubの報告として「過去数か月で、数十億リクエストにわたり新規処理が必要なプロンプトトークンの割合を50%以上削減」。 |
※ 「キャッシュを壊さずにエフォートを変えられる」は、エージェントのループで難易度に応じて推論の深さを変える運用と相性が良い変更です。仕組みの基礎は用語説明「プロンプトキャッシュ」で解説しています。
出典: OpenAI「Introducing GPT-6 Sol and Luna」(Improving caching節)
GPT-6ファミリー内での位置づけ
| モデル | 料金(入力/出力・1M) | 役割 | 当サイト |
|---|---|---|---|
| GPT-6 Astra | $10 / $50 | OpenAIは「continues to be our best model across the board」「Choose it when you want the best results and an uncompromising experience」と説明。最上位。 | 個別ページあり |
| GPT-6 Sol | $2 / $10 | 難しい業務タスクを、より高い利用上限と低コストで。SolはGPT-6世代の中位。 | 本ページ |
| GPT-6 Luna | $0.10 / $0.50 | 最安の実用ティア。Free / Goユーザーはデスクトップアプリで利用可。 | 本ページ |
| GPT-5.6 Sol | $4 / $20 | 前世代。GPT-6 Solに置き換え。旧版。 | 個別ページあり(旧版注記済み) |
| GPT-5.6 Luna | $0.20 / $1.20 | 前世代。GPT-6 Lunaに置き換え。旧版。 | 個別ページあり(旧版注記済み) |
※ Astraの料金は当サイト料金比較ページで確認した $10/$50(キャッシュ読込 $1.00)です。Astraは今回の値下げ対象ではなく、GPT-6世代の中で最上位に位置づけられています。
アラインメントと文体
- GPT-5.6版より改善: 公式は「both Sol and Luna show improvements over their GPT-5.6 counterparts, including lower rates of misleading claims about their coding work」と説明。ただし「The evaluations below deliberately test challenging situations and do not measure failure rates in typical use」と注記されています。
- system card: 詳細は deploymentsafety.openai.com/gpt-6-astra に掲載(AstraのカードがGPT-6世代の安全評価を兼ねる構成)。
- 文体はAstraから継承: 「more clarity, less jargon, fewer odd turns of phrase, fewer low-value details, and slightly shorter answers overall without losing substance」=技術・コーディング会話での改善が大きいとしています。
出典: OpenAI「Introducing GPT-6 Sol and Luna」(Continuing to improve alignment節・Collaboration style節)
選び方の目安
- コストを最優先する量産処理: GPT-6 Luna(入力$0.10 / 出力$0.50)は、価格帯として低価格モデルの最前列です。料金比較で他社の同価格帯と並べて確認してください。
- 難しい業務タスクを回す: GPT-6 Sol。AutomationBenchではClaude Opus 5の約1/11のタスク単価で上回ったとOpenAIは主張しています。
- Lunaは「弱い」わけではない: DeepSWE 1.1でLuna (max) 66.6%は、Claude Opus 5 および Fable 5 の mediumエフォート相当。エフォートを上げれば実力が出るタイプです。
- 旧世代ページとの混同注意: 検索から来た場合は、URLが
/gpt-5-6-sol/・/gpt-5-6-luna/(旧)か/gpt-6-sol-luna/(新)かを必ず確認してください。旧世代ページは当時の価格のまま保存しています。 - 上位の精度が必要なら: GPT-6 Astra が最上位です。GPT-6 Astra のページも参照してください。
⚠️ 免責事項
- 料金・仕様はOpenAI公式発表およびOpenAI APIドキュメントを2026-09-25に確認したものです。予告なく変動します。契約前に必ず公式ページをご確認ください。
- ベンチマーク数値はOpenAIの公表値で、競合モデルの数値は各社の公開レポートからの引用です。また一部のClaude比較値は「Fable 5」のもので、Fable 5.1とは別世代です。条件が同一とは限らず、実際の使用感を保証するものではありません。
- 本サイトは情報提供目的であり、特定プロバイダーの推薦・代理ではありません。