跳至正文
replicate / audio

realtime-tts-2

Most expressive text-to-speech model from Inworld, with natural-language steering, real-time latency, and multilingual support across 100+ languages.

inworld/realtime-tts-2

来源检查时间: 2026-09-13 · 供应商端点记录,不代表独立效果评测。

价格与适用条件。

以下是来源观察记录,不能直接相互替代。Token、计算耗时、输出时长与按次价格不能仅比较数值;输入和可选功能可能另行计费。

来源标价计费单位条件依据
$0.025 USDinput characteror 40,000 characters for $1公开结构化来源

来源提供的参数

languageLanguage of the input text. Use 'auto' to let the model detect the language. Supported production languages: English (en), Chinese (zh), Japanese (ja), Korean (ko), Russian (ru), Italian (it), Spanish (es), Portuguese (pt), French (fr), German (de), Polish (pl), Dutch (nl), Hindi (hi), Hebrew (he), Arabic (ar). · Options: auto, en, zh, ja, ko, ru, it, es, pt, fr, de, pl, nl, hi, he, ar · Default: auto
voice_idThe voice to use. Use a preset voice name (e.g. 'Ashley', 'Dennis', 'Alex', 'Darlene') or a custom cloned voice ID. · Default: Ashley
sample_rateAudio sample rate in Hz. · Options: 8000, 16000, 22050, 24000, 32000, 44100, 48000 · Default: 48000
audio_formatOutput audio format. · Options: mp3, wav, ogg_opus, flac · Default: mp3

未列出的参数、地区可用性、最低充值、税费和商用条款仍待核实,购买前请核对精确端点。

realtime-tts-2 · replicate API 价格 · Vidily