elevenlabs/eleven-v3/timing
ElevenLabs Eleven-V3 Timing converts text to natural speech and returns alignment metadata—character/word timestamps in JSON—for precise subtitles, karaoke effects, and lip-sync. Supports voice_id, similarity/stability, and optional Speaker Boost. Priced at $0.10 per 1,000 characters. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
elevenlabs/eleven-v3/timing
Source checked: 2026-09-13 · Provider endpoint, not an independently benchmarked model.
Pricing, with its conditions.
These are source observations, not interchangeable quotes. Token, compute-time, output-duration and per-request prices cannot be compared by the number alone. Input charges and optional features may add to the total.
| Source rate | Billing unit | Conditions | Evidence |
|---|---|---|---|
| 0.10 USD | source-listed starting price | Displayed catalog starting price only. Final charge depends on duration, resolution, outputs and references. Not a fixed per-second or per-image quote. | Public structured source |
Supported settings from the source
| category | text-to-audio |
|---|
Unlisted settings, region availability, deposits, taxes and commercial terms remain unverified. Confirm the exact endpoint before purchase.