テキスト読み上げ (Bytes)
Stream audio from a complete transcript
承認
Cartesia API key (sk_car_...). Get one at play.cartesia.ai/keys.
ヘッダー
API version header.
2026-08-14 "2026-08-14"
ボディ
The ID of the voice.
"db6b0ed5-d5d3-463d-ae85-518a07d3c2b4"
- WAVOutputFormat
- MP3OutputFormat
- RAWOutputFormat
The transcript's language or locale (for example en or en-GB). language and locale accept the same values. Set one or the other, never both; setting both returns a 400 error. See supported codes.
The transcript's language or locale (for example en or en-GB). locale and language accept the same values. Set one or the other, never both; setting both returns a 400 error. See supported codes.
Usually unnecessary: Cartesia picks the closest accent the voice supports for the requested language or locale. Set it only to make a multilingual voice sound accented (e.g. speak English with a French accent). Must come from the voice's Get Voice accents field. Learn more here.
Text normalization. auto (default) runs the locale-aware normalizer, off skips it, or pass a language or locale code (for example en or en-IN) to pin the normalizer independently of the generation language. See Text Normalization.
The ID of a pronunciation dictionary to use for the generation. Pronunciation dictionaries are supported by sonic-3 models and newer.
Configure the various attributes of the generated speech. Available on sonic-3 and newer models; not available on earlier models.
See Volume, Speed, and Emotion for a guide on this option.
レスポンス
Audio bytes
The response is of type file.