新規 無料登録、10回の呼び出しを進呈。最大 $1、カード不要。

GPT-6 Astra vs GPT-6.1 Sol

vs

いつ、どちらを使うか

これらはOpenAIのGPT-6ラインにおける2つの階層です。gpt-6-astraは最上位層であり、OpenAIはgpt-6.1-solを複雑なコーディング、コンピュータ操作、専門的作業向けの低コストなオプションとして位置付けており、ご自身のタスクでAstraと比較検討するためのものです。どちらもテキストおよび画像を入力とし、1050000トークンのコンテキストと128000の最大出力を備えているため、価格が判断基準となります。gpt-6-astraは入力が$10、出力が$50であり、$2および$10に対してどちらも5x高く、キャッシュ読み取りは$0.1に対し$1です。基本的にはgpt-6.1-solを使用し、評価によってアップグレードする価値があると示された場合にgpt-6-astraへ引き上げることとし、思考をオフにできるのはgpt-6-astraのみである点に注意してください。

ベンチマーク

GPT-6.1 Sol:ベンダーはベンチマークのスコアを公表していません。

平均超え他モデル以上GPT-6 Astra28 / 2921 / 29
GPT-6 Astra GPT-6.1 Sol 測定された他モデル 測定対象の平均 ★ 他モデルに上回られていない
DeepSWE 1.1
74.1%
N/A
GeneBench Pro
他モデルに上回られていない 37.1%
N/A
OSWorld 2.0 offline set, partial
他モデルに上回られていない 72.6%
N/A
ExploitBench (Cap%)
他モデルに上回られていない 100%
N/A
HealthBench Professional
他モデルに上回られていない 63.4%
N/A
Terminal-Bench-Science 0.1
他モデルに上回られていない 64.6%
N/A
GPQA Diamond
他モデルに上回られていない 96%
N/A
BrowseComp
他モデルに上回られていない 91.5%
N/A
OpenScore String Quartets (1 - OMR-NED) 0.19-0.84
0.84
N/A

ベンダー公表: Alibaba (Qwen) Anthropic ByteDance DeepSeek Google MiniMax Moonshot OpenAI Tencent Z.ai

料金

GPT-6 Astra GPT-6.1 Sol Δ
入力 / 1Mトークン $10 $2 5×
出力 / 1Mトークン $50 $10 5×
キャッシュ読み取り / 1Mトークン $1 $0.1 10×
キャッシュ書き込み 別途料金なし 別途料金なし -

ビルド時のライブカタログの料金です。各モデルのページには現在の料金カードが記載されています。

位置付け — この課金単位におけるすべての76個のチャットモデル全体の1Mトークンあたりの入力料金 (対数スケール)

GPT-6 Astra · $10 GPT-6.1 Sol · $2
$0.05 · Qwen3 VL Flash $30 · GPT-5.4 Pro

機能

GPT-6 Astra GPT-6.1 Sol
ツール使用 あり あり
思考コントロール 設定可能 常時オン
構造化出力 あり あり
プロンプトキャッシング 暗黙的 (自動) 暗黙的 (自動)
キャッシュ有効期間 5-10m, up to 1h 5-10m, up to 1h
最小キャッシュプレフィックス 1024 トークン 1024 トークン

仕様

GPT-6 Astra GPT-6.1 Sol
入力モダリティ テキスト 画像 テキスト 画像
出力モダリティ テキスト テキスト
リリース 2026-09-03 2026-09-29
知識のカットオフ 2026-04 2026-04
コンテキストウィンドウ 1.1M 1.1M
最大出力 128K 128K
思考パラメータ reasoning.effort reasoning.effort
指定可能な値
reasoning.effort
  • none
  • low
  • medium
  • high
  • xhigh
  • max
reasoning.effort
  • low
  • medium
  • high
  • xhigh
  • max
デフォルト medium medium

仕様は各ベンダーのドキュメントから転記されています。ベンダーが公開していない行は推測せず、省略しています。 すべての出典: GPT-6 Astra · GPT-6.1 Sol

1つのプロンプト、両方のモデル — ゲートウェイ経由で測定

プロンプト Explain to a CFO, in exactly three sentences, why our LLM bill doubled after we enabled reasoning. チェック ちょうど 3 文

GPT-6 Astra パス · 3 sentences

Enabling reasoning makes the model do more work per request, often generating internal reasoning tokens that are billed even though they do not appear in the final answer. If traffic, model choice, and pricing stayed the same, those extra tokens likely explain why our bill doubled—not longer answers or more users. We should verify the usage breakdown, then reserve higher reasoning levels for tasks where improved results justify the added cost.

出力 283 tok (+196 思考) レイテンシ 17.2 s

GPT-6.1 Sol パス · 3 sentences

Enabling reasoning makes the model do more computational work before answering, and many providers bill the resulting reasoning tokens even though users never see them. That means the same number of requests can cost twice as much because each request consumes more billable tokens, even if the visible answers stay the same length. We should confirm the increase in our usage breakdown and reserve higher reasoning levels for tasks where the improvement justifies the added cost.

出力 144 tok (+54 思考) レイテンシ 6.6 s

指示追従(ちょうど三文 — 数えられる)、読み手に合わせた調整(CFO 向けの語り口)、そして下のトークン計から見える隠れた思考の課金ギャップ。

プロンプト This function is supposed to return the longest run of consecutive equal items, but callers report it is off by one on single-element inputs and misses runs that end at the last index. Fix it and explain the bug in one paragraph. チェック 修正がテストを通る

GPT-6 Astra パス · 8/8 cases

```python def longest_run(xs): if not xs: return 0 best = cur = 1 for i in range(1, len(xs)): if xs[i] == xs[i - 1]: cur += 1 else: cur = 1 best = max(best, cur) return best ``` The original code counts equal adjacent pairs rather than items, so every nonempty run is undercounted by one: a single-element input returns `0`, and a run of three equal items returns `2`. Starting and resetting `cur` to `1` counts the first item of each run, while the empty-input check preserves a result of `0`. Updating `best` after each iteration includes runs ending at the last index without needing a final flush; the original also updated during trailing runs, but undercounted them.

出力 321 tok (+131 思考) レイテンシ 13.5 s

GPT-6.1 Sol パス · 8/8 cases

```python def longest_run(xs): if not xs: return 0 best = cur = 1 for i in range(1, len(xs)): if xs[i] == xs[i - 1]: cur += 1 else: cur = 1 best = max(best, cur) return best ``` The original function counts equal adjacent pairs rather than items in a run: a run of length `n` has only `n - 1` equal pairs, so a single-element input returns `0`, and a run ending at the last index is also undercounted. It does process the last index, so the issue is not a missing final check. Initializing each nonempty run to `1` and updating `best` after every iteration correctly counts items, including singleton and trailing runs; empty inputs still return `0`.

出力 189 tok レイテンシ 6.3 s

修正が実際に正しいか(実行可能か)、説明の密度、そして範囲の限られたタスクでのトークン効率。

プロンプト Extract a JSON object with fields {company, quarter, revenue_usd_m, growth_pct, guidance} from this text. Use null for anything not stated; add no other fields. チェック 妥当な JSON、スキーマ厳密

GPT-6 Astra パス · 5/5 fields, guidance "Formal guidance is withheld pending the "

{ "company": "Meridian Robotics", "quarter": "September quarter", "revenue_usd_m": null, "growth_pct": "eleven and change", "guidance": "Formal guidance is withheld pending the Q4 close." }

出力 329 tok (+267 思考) レイテンシ 16.9 s

GPT-6.1 Sol パス · 5/5 fields, guidance "Withheld pending the Q4 close"

{ "company": "Meridian Robotics", "quarter": "September quarter", "revenue_usd_m": null, "growth_pct": "eleven and change", "guidance": "Withheld pending the Q4 close" }

出力 345 tok (+285 思考) レイテンシ 10.2 s

スキーマ遵守(フィールドをでっち上げない)、ハルシネーション圧(guidance は明示的に保留されている)、そして構造化出力の経路の違い。

プロンプト Write a 120-word product announcement for an API gateway feature that caches prompts across providers. Forbidden words: "seamless", "unlock", "game-changer", "revolutionize", "empower". Exactly one sentence must be a question. チェック 120 語、禁止語 0

GPT-6 Astra パス · 120 words, 0 banned, 1 question

Introducing Cross Provider Prompt Cache, a new API gateway feature that reuses eligible prompt content across supported AI providers. Why pay to process the same context every time? Configure caching once at the gateway, then route requests between models while keeping shared instructions, reference material, and conversation prefixes ready for reuse. Caching controls let teams set expiration windows, isolate tenants, and exclude sensitive content. Cache analytics show hit rates, estimated savings, and latency trends, helping developers tune performance with confidence. Existing routing policies continue to work, so adoption fits your current architecture. Start with a single application, measure the results, and expand as needed. Available today in the dashboard and API, with documentation and examples to guide your first deployment.

出力 665 tok (+516 思考) レイテンシ 19.3 s

GPT-6.1 Sol パス · 120 words, 0 banned, 1 question

Introducing Cross-Provider Prompt Cache, a new API gateway feature that stores reusable prompts and manages caching across your supported AI providers. Why rebuild the same context every time your application switches models? With one configuration, teams can reuse shared instructions, standardize cache policies, and reduce repeated prompt processing wherever provider caching is available. The gateway handles provider-specific requirements while giving you clear visibility into cache hits, usage, and estimated savings. Set expiration windows, isolate cached content by project, and invalidate entries when prompts change. Your existing routing logic stays intact, so you can compare models without rebuilding your caching workflow. […]

出力 588 tok (+435 思考) レイテンシ 13.9 s

制約の遵守(語数の上限、禁止語リスト、唯一の疑問文)、文体の指紋、そして長さの制御。

1行で切り替え

以下のすべてのタブには両方のIDが含まれています — 変更箇所はハイライトされた2行のみです。エンドポイント、キー、リクエスト形式はすべて同じです。

from openai import OpenAI

client = OpenAI(
    base_url="https://synthorai.io/v1",
    api_key="sk-syn-...",
)

resp = client.chat.completions.create(
    model="gpt-6-astra",
    # model="gpt-6.1-sol",  # この行をアンコメントし、上の行をコメントアウトします
    messages=[{"role": "user", "content": "Summarize this diff"}],
    reasoning_effort="medium",
)
print(resp.choices[0].message.content)

API キーを取得 →

FAQ

GPT-6 Astra と GPT-6.1 Sol ではどちらが安いですか?

入力 / 1mトークン においては GPT-6.1 Sol の方が安価です($2 対 $10、5.0× の差)。他の行では逆になる可能性があります — 上の表には完全な情報が記載されており、実際のコストは組み合わせに依存します。

2つの統合を行わずに GPT-6 Astra と GPT-6.1 Sol のA/Bテストを実施できますか?

はい。両方とも1つのAPIキーで同じOpenAI互換エンドポイントを通じて提供されます — モデル文字列を1行変更するだけで切り替えられるため、トラフィックの一部をそれぞれにルーティングし、請求額を直接比較できます。

GPT-6 Astra と GPT-6.1 Sol はプロンプトキャッシングをサポートしていますか?

はい — どちらのモデルでもキャッシュ読み込みは入力レートよりも低く請求されるため、ウォームプレフィックスのワークロードは定価が示すよりも低コストになります。キャッシュ読み込みの正確な行は、上の料金表に記載されています。

関連する比較

当社の測定調査より