跳至正文
全部媒体模型/WaveSpeedAI/elevenlabs/eleven-v3/timing
wavespeed / audio

elevenlabs/eleven-v3/timing

ElevenLabs Eleven-V3 Timing converts text to natural speech and returns alignment metadata—character/word timestamps in JSON—for precise subtitles, karaoke effects, and lip-sync. Supports voice_id, similarity/stability, and optional Speaker Boost. Priced at $0.10 per 1,000 characters. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

elevenlabs/eleven-v3/timing

来源检查时间: 2026-09-13 · 供应商端点记录,不代表独立效果评测。

价格与适用条件。

以下是来源观察记录,不能直接相互替代。Token、计算耗时、输出时长与按次价格不能仅比较数值;输入和可选功能可能另行计费。

来源标价计费单位条件依据
0.10 USDsource-listed starting priceDisplayed catalog starting price only. Final charge depends on duration, resolution, outputs and references. Not a fixed per-second or per-image quote.公开结构化来源

来源提供的参数

categorytext-to-audio

未列出的参数、地区可用性、最低充值、税费和商用条款仍待核实,购买前请核对精确端点。