新帳號 免費註冊,送 10 次呼叫,最高 $1,免綁卡。

BytePlus Seed TTS 2.0 vs TTS-1 HD

vs

什麼情況選哪一個

就已知規格而言,這兩者是可以互換的:seed-tts-2.0 和 tts-1-hd 皆接收文字輸入並回傳音訊,皆具備語音能力,最高皆限制在 4096-token 的上下文,並且皆對每百萬輸入 token 收費 $30。由於費率表和單位完全相符,無論哪種選擇都沒有成本或上下文的爭議——請根據供應商偏好、適合你產品的語音輸出,或是你已經整合了哪家供應商來進行選擇。透過這兩者分別執行一段你自己腳本的簡短樣本,並保留聽起來合適的那一個。

定價

BytePlus Seed TTS 2.0 TTS-1 HD Δ
每 1M 字元 $30 $30 =

費率取自網站建置時的即時目錄;各模型頁面都列有最新的價目。

兩者的相對位置:每 1M 字元的價格,涵蓋同一計費單位下全部 7 個文字轉語音模型(對數尺度)

功能

BytePlus Seed TTS 2.0 TTS-1 HD
串流 是 是
SSML unsupported undocumented
計費單位 character character

規格

BytePlus Seed TTS 2.0 TTS-1 HD
輸入模態 文字 文字
輸出模態 音訊 音訊
發布日期 2025-10-16 2023-11-06
請求限制 - 4096 characters
聲音

TTS 2.0 voices carry *_uranus_bigtts speaker IDs and are listed by scenario (General, Entertainment, Education, Dubbing, AudioBook, RolePlay, CustomerService) on the Voice List page

per-voice emotion via audio_params.emotion with emotion_scale 1 to 5 (default 4), speech_rate and loudness_rate in [-50, 100], and post_process.pitch in [-12, 12].

alloy, ash, coral, echo, fable, onyx, nova, sage and shimmer, a smaller set than the full 13-voice TTS list, and voices are currently optimized for English.
語言 English, Chinese, Japanese, German, French, Mexican Spanish, Indonesian Bahasa, Brazilian Portuguese, Italian and Korean Generally follows the Whisper model’s language support: Afrikaans, Arabic, Armenian, Azerbaijani, Belarusian, Bosnian, Bulgarian, Catalan, Chinese, Croatian, Czech, Danish, Dutch, English, Estonian, Finnish, French, Galician, German, Greek, Hebrew, Hindi, Hungarian, Icelandic, Indonesian, Italian, Japanese, Kannada, Kazakh, Korean, Latvian, Lithuanian, Macedonian, Malay, Marathi, Maori, Nepali, Norwegian, Persian, Polish, Portuguese, Romanian, Russian, Serbian, Slovak, Slovenian, Spanish, Swahili, Swedish, Tagalog, Tamil, Thai, Turkish, Ukrainian, Urdu, Vietnamese and Welsh, despite voices being optimized for English
聲音控制 context_texts and section_id are TTS 2.0 only: a plain-language instruction steers rate, emotion, volume and style (only the first list value takes effect, and its text is not billed), and section_id links up to 30 rounds or 10 minutes of earlier synthesis as historical context. none
限制

Integration via uni-directional streaming HTTP, uni- or bi-directional streaming WebSocket and online SDK

sampling rate 24K/16K/8K on the uni-directional streaming and non-streaming interfaces, 48K/24K/16K/8K on the bi-directional interface

output PCM, OGG_OPUS or MP3 (WAV is accepted but returns multiple headers when streaming)

SSML is not supported

Speech generation endpoint (v1/audio/speech) only, with Chat Completions, Responses, Realtime, Batch and Fine-tuning all listed as not supported

input text is capped at 4096 characters

output mp3 (default), opus, aac, flac, wav or pcm (raw 24 kHz 16-bit signed little-endian samples)

realtime playback via chunked transfer encoding

規格摘錄自各供應商的文件;供應商沒有公布的項目就直接略過,不自行推測。 完整來源: BytePlus Seed TTS 2.0 · TTS-1 HD

改一行程式碼就能在兩者之間切換

下面每個頁籤都列了這兩個模型 ID,要改的只有醒目標示的那兩行。端點、金鑰和請求格式都不變。

from openai import OpenAI

client = OpenAI(
    base_url="https://synthorai.io/v1",
    api_key="sk-syn-...",
)

resp = client.audio.transcriptions.create(
    model="seed-tts-2.0",
    # model="tts-1-hd",  # 取消這一行的註解,並把上一行註解掉
    file=open("meeting.mp3", "rb"),
    language="en",
)
print(resp.text)

取得 API 金鑰 →

常見問題

BytePlus Seed TTS 2.0 和 TTS-1 HD 哪個比較便宜?

兩者的「每 1M 字元」相同($30),這一項分不出高下,請看下方的規格與功能。

可以只串接一次,就對 BytePlus Seed TTS 2.0 和 TTS-1 HD 做 A/B 測試嗎?

可以。兩個模型都走同一個 OpenAI 相容端點,用的也是同一把 API 金鑰,切換時只要改一行裡的模型名稱字串。你可以把一部分流量分別導到兩邊,再直接比較帳單。

文字轉語音是如何計費的?

依輸入文字的字元數計費,每次請求的字元上限列在規格表裡。兩個模型都一樣,稿子太長就得拆成多次請求。

相關比較