Skip to main content
Gradium TTS is a real-time streaming model with expressive, natural speech across English, Spanish, French, German, and Portuguese. Voices are referenced by ID. Pass the voice_id from the tables below in your synthesis request.
The default output format is wav. Provide a custom voice ID in the same voice_id field to use a cloned voice.

English

Spanish

French

German

Portuguese