Skip to content
All media models/Replicate/realtime-tts-1.5-max
replicate / audio

realtime-tts-1.5-max

Highest-quality realtime text-to-speech with <200ms latency, emotion control, and 15-language support

inworld/realtime-tts-1.5-max

Source checked: 2026-09-13 · Provider endpoint, not an independently benchmarked model.

Pricing, with its conditions.

These are source observations, not interchangeable quotes. Token, compute-time, output-duration and per-request prices cannot be compared by the number alone. Input charges and optional features may add to the total.

Source rateBilling unitConditionsEvidence
$0.035 USDinput characteror around 28,571 characters for $1Public structured source

Supported settings from the source

voice_idThe voice to use. Use a preset voice name (e.g. 'Ashley', 'Dennis', 'Alex') or a custom cloned voice ID. · Default: Ashley
sample_rateAudio sample rate in Hz. · Options: 8000, 16000, 22050, 24000, 32000, 44100, 48000 · Default: 48000
audio_formatOutput audio format. · Options: mp3, wav, ogg_opus, flac · Default: mp3

Unlisted settings, region availability, deposits, taxes and commercial terms remain unverified. Confirm the exact endpoint before purchase.