wavespeed / audio
elevenlabs/eleven-v3/timing
ElevenLabs Eleven-V3 Timing converts text to natural speech and returns alignment metadata—character/word timestamps in JSON—for precise subtitles, karaoke effects, and lip-sync. Supports voice_id, similarity/stability, and optional Speaker Boost. Priced at $0.10 per 1,000 characters. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
elevenlabs/eleven-v3/timing
来源检查时间: 2026-09-13 · 供应商端点记录,不代表独立效果评测。
价格与适用条件。
以下是来源观察记录,不能直接相互替代。Token、计算耗时、输出时长与按次价格不能仅比较数值;输入和可选功能可能另行计费。
| 来源标价 | 计费单位 | 条件 | 依据 |
|---|---|---|---|
| 0.10 USD | source-listed starting price | Displayed catalog starting price only. Final charge depends on duration, resolution, outputs and references. Not a fixed per-second or per-image quote. | 公开结构化来源 |
来源提供的参数
| category | text-to-audio |
|---|
未列出的参数、地区可用性、最低充值、税费和商用条款仍待核实,购买前请核对精确端点。