GPT-5.4 是 OpenAI 面向編碼與專業工作的前沿模型,在模型頁面上被定位為「一款面向編碼與專業工作、更實惠的模型」。
- 輸入
- 文字 影像 $2.5/M
- 輸出
- 文字 $15/M
- 快取讀取
- $1.25/M
- 上下文
- 922K
- 相較 GPT-4o
- 便宜約 50%
- 知識截止
- 2025-08
Benchmark 成績
廠商公布: Alibaba (Qwen) Anthropic ByteDance DeepSeek Google MiniMax Moonshot OpenAI Tencent Z.ai
價格在同類中的位置
價格在 60 個同類模型中的位置
這條線顯示該模型的價格,在 Synthorai 上同類模型裡處於什麼位置。兩端標出了最便宜和最貴的那個。這裡是基礎價,批次、區域與快取寫入的折扣見價格頁。
規格與限制
Token
| 上下文視窗(廠商規格) | 1,050,000 |
|---|---|
| 最大輸出(廠商規格) | 128,000 |
| 知識截止 | 2025-08 |
提示詞快取
| 快取方式 | 自動 |
|---|---|
| 最低前綴 | 1,024 |
| 存活時間 | 5–10 分鐘,最長 1 小時 |
思考
| 廠商參數 | reasoning.effort |
|---|---|
| 可選值 | none · low · medium · high · xhigh |
| 預設值 | none 請求未指定時生效 |
| 可關閉 | 支援 |
| 思考行為 | 支援交錯思考:思考之前與思考之間都可以有可見輸出,工具呼叫之間也可以思考。 |
| 參數 | reasoning_effort |
| 取值 | minimal · low · medium · high 閘道側參數面——以上方廠商映射為準 |
模型
| 模態 | 文字 + 影像 → 文字 |
|---|
- 1.05M token 上下文 / 128k 最大輸出
- 輸入超過 272k token 的提示詞按 2 倍輸入 / 1.5 倍輸出計費
- 推理力度 none 至 xhigh
單一提示詞 — 經由閘道測量
GPT-5.4 通過 · 3 sentences
Enabling reasoning makes the model generate and evaluate additional internal steps before producing an answer, which increases the total tokens processed and therefore the cost per request. It also tends to increase latency and can trigger more tool calls, retries, or longer context retention, all of which compound usage-based charges. So even if user traffic stayed flat, the average cost per interaction rose materially once reasoning was turned on.
輸出 82 tok 延遲 1.9 s
指令遵循(恰好三句,可數)、受眾適配(面向 CFO 的語氣),以及下方 token 計量所暴露的隱藏思考計費缺口。
GPT-5.4 通過 · 8/8 cases
```python def longest_run(xs): if not xs: return 0 best = 1 cur = 1 for i in range(1, len(xs)): if xs[i] == xs[i - 1]: cur += 1 else: best = max(best, cur) cur = 1 best = max(best, cur) return best ``` The bug is that the original code was counting *matching transitions* between adjacent elements instead of the *length of the run itself*, so a single element started at `0` instead of `1`, making all run lengths off by one. It also only updated `best` inside the equality case, which means the final run could be lost if the loop ended before a change triggered any correction logic. […]
輸出 243 tok 延遲 3.0 s
修復是否真的正確(可執行)、解釋的資訊密度,以及在一個邊界明確的任務上的 token 效率。
GPT-5.4 通過 · 5/5 fields, guidance "withheld pending the Q4 close"
{"company":"Meridian Robotics","quarter":"September quarter","revenue_usd_m":412,"growth_pct":11,"guidance":"withheld pending the Q4 close"}
輸出 39 tok 延遲 1.3 s
schema 服從度(不臆造欄位)、幻覺壓力(guidance 明確被暫緩給出),以及結構化輸出路徑的差異。
GPT-5.4 通過 · 126 words, 0 banned, 1 question
Today we’re introducing Prompt Cache, a new API gateway feature that stores prompt results and reuses them across supported model providers. Teams can reduce latency, lower repeated compute costs, and keep application behavior more consistent when identical or near-identical prompts appear in production. Prompt Cache works at the gateway layer, so developers can adopt it without rewriting provider integrations or changing routing logic. Built-in controls let you set cache keys, expiration policies, and privacy rules for sensitive workloads. Need faster responses without duplicating optimization work for each vendor? With shared observability, you can track hit rates, savings, and cache performance in one place. […]
輸出 154 tok 延遲 2.9 s
約束服從度(字數預算、禁用詞表、唯一的那句問句)、文風指紋,以及長度控制。
30 秒用上 GPT-5.4
OpenAI 相容:換掉 base_url,SDK 不用改。POST /v1/chat/completions
from openai import OpenAI
client = OpenAI(
base_url="https://synthorai.io/v1",
api_key="sk-syn-...",
)
resp = client.chat.completions.create(
model="gpt-5.4",
messages=[{"role": "user", "content": "Summarize this diff"}],
reasoning_effort="medium",
)
print(resp.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://synthorai.io/v1",
apiKey: "sk-syn-...",
});
const resp = await client.chat.completions.create({
model: "gpt-5.4",
messages: [{ role: "user", content: "Summarize this diff" }],
reasoning_effort: "medium",
});
console.log(resp.choices[0].message.content);curl https://synthorai.io/v1/chat/completions \
-H "Authorization: Bearer sk-syn-..." \
-H "Content-Type: application/json" \
-d '{
"model": "gpt-5.4",
"messages": [{"role": "user", "content": "Hello"}],
"reasoning_effort": "medium"
}'package main
import (
"context"
"fmt"
"github.com/openai/openai-go/v3"
"github.com/openai/openai-go/v3/option"
)
func main() {
client := openai.NewClient(
option.WithBaseURL("https://synthorai.io/v1"),
option.WithAPIKey("sk-syn-..."),
)
resp, _ := client.Chat.Completions.New(context.TODO(), openai.ChatCompletionNewParams{
Model: "gpt-5.4",
Messages: []openai.ChatCompletionMessageParamUnion{
openai.UserMessage("Summarize this diff"),
},
ReasoningEffort: openai.ReasoningEffortMedium,
})
fmt.Println(resp.Choices[0].Message.Content)
}import com.openai.client.OpenAIClient;
import com.openai.client.okhttp.OpenAIOkHttpClient;
import com.openai.models.chat.completions.*;
import com.openai.models.ReasoningEffort;
OpenAIClient client = OpenAIOkHttpClient.builder()
.baseUrl("https://synthorai.io/v1")
.apiKey("sk-syn-...")
.build();
ChatCompletion resp = client.chat().completions().create(
ChatCompletionCreateParams.builder()
.model("gpt-5.4")
.addUserMessage("Summarize this diff")
.reasoningEffort(ReasoningEffort.MEDIUM)
.build());
System.out.println(resp.choices().get(0).message().content().orElse(""));關於 GPT-5.4
- 它把這個家族的上下文視窗擴展到 1,050,000 token(廠商規格),最大輸出 128K token,支援影像輸入、從 none 到 xhigh 可設定的推理力度、折扣快取輸入,以及包含網頁搜尋與程式碼直譯器在內的完整託管工具集,知識截止為 2025 年 8 月。
- 推理力度預設為 none,這讓延遲維持在低點,但也意味著需要斟酌的工作必須主動要求。
- 發布時 OpenAI 指出這款模型有四項新能力:1M 級的上下文視窗、工具搜尋、內建的電腦使用,以及面向長時間執行會話的壓縮(compaction)。
- 除此之外,工具清單還包含檔案搜尋、影像生成、程式碼直譯器、託管 shell、apply patch、skills 與 MCP,而 Chat Completions、Responses 與 Batch 全部受支援。
- 有一項計費機制值得預先規劃:輸入超過 272K token 的 prompt,整個請求會按 2 倍輸入與 1.5 倍輸出計費,也就是說,這筆附加費由輸入大小觸發,卻同時抬高了輸出費率。
- 當旗艦級能力必須兼顧生產經濟性時,它就是預設之選。
- Synthorai 使用者透過閘道器的 OpenAI 相容 chat completions 端點使用 GPT-5.4。
常見問題
GPT-5.4 API 可以免費試用嗎?
可以,新帳號可獲得 10 次試用呼叫和最高 $1 的免費額度,無需信用卡。以輸入 $2.5/M 計算,光是這筆額度就足以對 GPT-5.4 發出約 49 次 ~8K token 的請求。
GPT-5.4 最擅長什麼?
1,050,000 token 上下文視窗、更實惠的旗艦級選項、含網頁搜尋的完整託管工具集。完整能力請見「關於」一節,內容取自廠商官方發布說明。
GPT-5.4 的價格是多少?
在 Synthorai 上,GPT-5.4 輸入 $2.5/百萬 token、輸出 $15/百萬 token,即廠商牌價,無平台加價。快取命中的輸入 token 以 $1.25/M 計費。
GPT-5.4 支援提示詞快取(prompt caching)嗎?
支援,且全自動:由 OpenAI 供應的提示詞會自動快取,無需改程式碼。快取命中的輸入 token 以 $1.25/M 計費(未命中 $2.5/M);提示詞需有 1,024 個 token 以上的穩定前綴才能命中快取(TTL 5–10 分鐘,最長 1 小時)。 提示詞快取指南 →
如何開通 GPT-5.4?
把現有 OpenAI SDK 的 base_url 指向 "https://synthorai.io/v1",model 設為 "gpt-5.4" 即可。一組 API key 通用閘道上的所有模型。
GPT-5.4 的知識截止日期是什麼時候?
GPT-5.4 的知識截止日期為 2025-08,依據廠商官方文件(資料核驗於 2026-07-09)。
相關模型
對比
本頁每個值都轉錄自廠商自己的文件(連結見上),並帶有核對日期。價格在全目錄範圍內比較;各廠商定義不同的規格值,只說明差異而不作圖表對比。此處沒有任何由我們測量的資料,也不做評分。