Audio APIs
Convert text to natural-sounding speech using Melo TTS. Generate high-quality audio in multiple languages with customizable speed and speaker options.Text-to-Speech
Endpoint
Basic Example
- Python
- cURL
Request Parameters
Required Parameters
Optional Parameters
Response Format
The API returns a JSON object containing base64-encoded MP3 audio:Decoding the Response
Supported Languages
Melo TTS supports 6 languages with various speaker options:English has multiple speaker variants for different accents: US (American), BR (British), INDIA (Indian), and AU (Australian).
Using Different Languages and Speakers
- Python
- cURL
Multilingual Example
Speed Control
Adjust thespeed parameter to control how fast the speech is generated:
Example
Pricing
Rate: $5.00 per 1 million charactersThere are no character limits per request. You are billed based on the total characters processed.
Use Cases
- Voice assistants: Add natural speech to chatbots and virtual assistants
- Audiobook generation: Convert written content to audio format
- Accessibility: Make content accessible for visually impaired users
- Video narration: Generate voiceovers for videos and presentations
- Language learning: Create pronunciation examples in multiple languages
- Notification systems: Generate audio alerts and announcements
Next Steps
Text APIs
Generate text with large language models
Vision Language Models
Analyze images with multimodal AI
Image APIs
Generate images from text prompts

