MiniMax M3 vs Qwen3.8 Max
MiniMax M3 已從我們的目錄下架。下方是它最後公布的費率;它已經無法呼叫,這組比較裡的另一個模型則仍可呼叫。
什麼情況選哪一個
minimax-m3 在價目表的每一行都更便宜,輸入 $0.3、輸出 $1.2,對比 qwen3.8-max 的 $2 和 $6——約低 6.7 倍和 5 倍——而且它還在文字和影像之外接受視訊,帶 1000000 token 脈絡和最多 524288 token 輸出,並允許關閉思考。qwen3.8-max 更貴,輸出上限 131072 token,脈絡 983616 token,所以當你想要阿里較新的 2026 年 8 月模型、在文字和影像工作上帶明確視覺標誌時用它。大批量、長輸出或視訊輸入的任務,minimax-m3 是划算的選擇。
Benchmark 成績
MiniMax M3:供應商沒有公布 benchmark 成績。
供應商公布: Alibaba (Qwen) Anthropic ByteDance DeepSeek Google MiniMax Moonshot OpenAI Tencent Z.ai
定價
| MiniMax M3 | Qwen3.8 Max | Δ | |
|---|---|---|---|
| 輸入 / 1M tokens | $0.3 | $2 | 0.15× |
| 輸出 / 1M tokens | $1.2 | $6 | 0.2× |
| 快取讀取 / 1M tokens | $0.06 | $0.25 | 0.24× |
| 快取寫入 | 不額外計費 | 1.25x | - |
費率取自網站建置時的即時目錄;各模型頁面都列有最新的價目。
兩者的相對位置:每 1M tokens 的輸入價格,涵蓋同一計費單位下全部 76 個聊天模型(對數尺度)
功能
| MiniMax M3 | Qwen3.8 Max | |
|---|---|---|
| 工具使用 | 是 | 是 |
| 思考控制 | 可設定 | 是,但供應商未公布調整參數 |
| 結構化輸出 | - | 是 |
| 提示詞快取 | 隱式(自動) | 隱式 + 顯式 |
| 快取存活時間 | 未公布 | explicit: 5m, reset on hit |
| 最小快取前綴 | 512 個 token | 1024 個 token |
規格
| MiniMax M3 | Qwen3.8 Max | |
|---|---|---|
| 輸入模態 | 文字 圖像 影片 | 文字 圖像 |
| 輸出模態 | 文字 | 文字 |
| 發布日期 | 2026-06-01 | 2026-08-03 |
| 上下文視窗 | 1M | 984K |
| 最大輸出 | 524K | 131K |
| 思考參數 |
| - |
| 可接受的值 | thinking.type
reasoning_split
| - |
| 預設值 | adaptive: thinking on, with the model deciding when extra reasoning helps | - |
規格摘錄自各供應商的文件;供應商沒有公布的項目就直接略過,不自行推測。 完整來源: MiniMax M3 · Qwen3.8 Max
改一行程式碼就能在兩者之間切換
下面每個頁籤都列了這兩個模型 ID,要改的只有醒目標示的那兩行。端點、金鑰和請求格式都不變。
from openai import OpenAI
client = OpenAI(
base_url="https://synthorai.io/v1",
api_key="sk-syn-...",
)
resp = client.chat.completions.create(
model="minimax-m3",
# model="qwen3.8-max", # 取消這一行的註解,並把上一行註解掉
messages=[{"role": "user", "content": "Summarize this diff"}],
reasoning_effort="medium",
)
print(resp.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://synthorai.io/v1",
apiKey: "sk-syn-...",
});
const resp = await client.chat.completions.create({
model: "minimax-m3",
// model: "qwen3.8-max", // 取消這一行的註解,並把上一行註解掉
messages: [{ role: "user", content: "Summarize this diff" }],
reasoning_effort: "medium",
});
console.log(resp.choices[0].message.content);curl https://synthorai.io/v1/chat/completions \
-H "Authorization: Bearer sk-syn-..." \
-H "Content-Type: application/json" \
-d '{
"model": "minimax-m3",
# "model": "qwen3.8-max", # 取消這一行的註解,並把上一行註解掉
"messages": [{"role": "user", "content": "Hello"}],
"reasoning_effort": "medium"
}'package main
import (
"context"
"fmt"
"github.com/openai/openai-go/v3"
"github.com/openai/openai-go/v3/option"
)
func main() {
client := openai.NewClient(
option.WithBaseURL("https://synthorai.io/v1"),
option.WithAPIKey("sk-syn-..."),
)
resp, _ := client.Chat.Completions.New(context.TODO(), openai.ChatCompletionNewParams{
Model: "minimax-m3",
// Model: "qwen3.8-max", // 取消這一行的註解,並把上一行註解掉
Messages: []openai.ChatCompletionMessageParamUnion{
openai.UserMessage("Summarize this diff"),
},
ReasoningEffort: openai.ReasoningEffortMedium,
})
fmt.Println(resp.Choices[0].Message.Content)
}import com.openai.client.OpenAIClient;
import com.openai.client.okhttp.OpenAIOkHttpClient;
import com.openai.models.chat.completions.*;
import com.openai.models.ReasoningEffort;
OpenAIClient client = OpenAIOkHttpClient.builder()
.baseUrl("https://synthorai.io/v1")
.apiKey("sk-syn-...")
.build();
ChatCompletion resp = client.chat().completions().create(
ChatCompletionCreateParams.builder()
.model("minimax-m3")
// .model("qwen3.8-max") // 取消這一行的註解,並把上一行註解掉
.addUserMessage("Summarize this diff")
.reasoningEffort(ReasoningEffort.MEDIUM)
.build());
System.out.println(resp.choices().get(0).message().content().orElse(""));常見問題
MiniMax M3 和 Qwen3.8 Max 哪個比較便宜?
以「輸入 / 1M tokens」來看,MiniMax M3 比較便宜($0.3 對 $2,相差 6.7×)。其他項目的結果可能相反,完整價目請看上表;實際成本要看你的用量組合。
可以只串接一次,就對 MiniMax M3 和 Qwen3.8 Max 做 A/B 測試嗎?
可以。兩個模型都走同一個 OpenAI 相容端點,用的也是同一把 API 金鑰,切換時只要改一行裡的模型名稱字串。你可以把一部分流量分別導到兩邊,再直接比較帳單。
MiniMax M3 與 Qwen3.8 Max 支援提示詞快取嗎?
支援。兩者的快取讀取費率都低於輸入費率,所以前綴已經進快取的工作負載,實際成本會比官網價算出來的低。確切的快取讀取價格請見上方定價表。