Convert text to speech as an MP3 file

Turn your text into natural speech with Sonic. Choose a voice, listen to your script, and download an MP3 for app playback, phone prompts, or narration.
242/500

Enter your text, choose a voice, and create your MP3.

Your script, ready to play

Check the voice with a short passage, then generate a file for your player. Keep the original script so you can revise the recording later.

Prepare the script

Write the text you want in the file. Spell out unfamiliar abbreviations and review names, dates, and numbers before generating a final recording.

Listen to a preview

Try your text in the preview and compare voices. Listen to the delivery before creating your MP3 file.

Save the MP3

Select Create MP3, then download your audio file. Listen once more in the player your audience will use.

Generate an MP3 through the API

Send your script to Sonic, then save the response as an MP3 for narration, lessons, or app playback.

Configure your request

Choose MP3 output and include your transcript, voice, and model in a Bytes API request. Send the request from your backend to keep your API key private.

Save the audio response

Check that the request succeeded before saving the response bytes as an MP3 file. Handle any API errors first, then play the saved audio to check the result.

MP3 output settings
output_format = {
    "container": "mp3",
    "sample_rate": 44100,
    "bit_rate": 128000,
}

Format
MP3
Sample rate
44.1 kHz
Bit rate
128 kbps

Choose a format for your destination

Start with what your player accepts. A phone integration and a downloadable narration file may need different settings.

MP3

For playback and sharing

Compressed audio for narration, app messages, and players that accept MP3. Choose the bit rate and sample rate in your request.

WAV

For an editing workflow

A file container for PCM audio. Useful when your editor or publishing workflow expects WAV rather than compressed MP3.

Raw audio

For an integration

Audio samples without a file header. Match the encoding and sample rate to your player or telephone system.

Put the file to work

Fixed phone prompts

Generate greetings and closed-hours messages from approved text. Confirm your phone system accepts MP3 before uploading; some systems require WAV or raw audio instead.

In-app audio

Save spoken instructions or short messages for playback in your application. Keep the source text alongside each file so you know which version users will hear.

Voice evaluation

Compare voices using the same test sentences. Review pronunciation and pacing before choosing a voice for a customer-facing agent.

Text-to-MP3 FAQs

Keep working with your voice

Compare voices for the next script, or learn how to generate speech from your application.

Find a voice for your script

Compare voices on a short passage before producing the final file. Use the voice generator to hear how each voice handles your words.

Try the voice generator

Generate audio from Python

Set up the Python SDK and make your first text-to-speech request. Start with the complete example, then choose the output format your project needs.

Build with Python

Get started today

Talk to an expert.

Connect with a member of our team and learn how Cartesia can help you build world-class voice experiences.

Contact Sales

Start building.

Access our models via API and bring a voice agent into production in minutes.

Try Cartesia