テキスト読み上げ (Bytes)
Stream audio from a complete transcript
承認
Cartesia API key (sk_car_...). Get one at play.cartesia.ai/keys.
ヘッダー
API version header.
2026-08-14 "2026-08-14"
ボディ
The ID of the voice.
"db6b0ed5-d5d3-463d-ae85-518a07d3c2b4"
- WAVOutputFormat
- MP3OutputFormat
- RAWOutputFormat
Prefer locale when you can. language only accepts base ISO codes like en. A request may set language or locale, never both.
en, fr, de, es, pt, zh, ja, hi, it, ko, nl, pl, ru, sv, tr, tl, bg, ro, ar, cs, el, fi, hr, ms, sk, da, ta, uk, hu, no, vi, bn, th, he, ka, id, te, gu, kn, ml, mr, pa, or, ur Prefer locale over language (for example en-GB). language only accepts base codes like en. locale also accepts regional codes like en-GB. Locale codes need Sonic 3.6+. A request may set language or locale, never both.
Sets how the voice sounds while speaking the language. Only needed for a multilingual voice when you want accented speech. Must be one of the voice's supported accents, based on the accents field from Get Voice.
arabic, khaleeji, middle-eastern-arabic, modern-standard-arabic Text normalization. auto (default) runs the locale-aware normalizer, off skips it, or pass a locale code (for example en-IN) to pin the normalizer independently of the generation language. See Text Normalization.
The ID of a pronunciation dictionary to use for the generation. Pronunciation dictionaries are supported by sonic-3 models and newer.
Configure the various attributes of the generated speech. Available on sonic-3 and newer models; not available on earlier models.
See Volume, Speed, and Emotion for a guide on this option.
レスポンス
Audio bytes
The response is of type file.