Text-to-speech studio for Mac

Studio-quality voices, made on your Mac

Type a script, pick a voice and hear it spoken in seconds. Clone your own voice, design a new one from a description, teach a voice to sound like you, and keep every take organised. Free for Apple silicon Macs, coming soon.

Coming soon for MacFree for Apple silicon Macs · macOS 15 or laterSee every feature
Avocado’s Generate page with a script, a model and voice picker, and recent takes listed beside itAvocado’s Generate page with a script, a model and voice picker, and recent takes listed beside it

39

open voice model families Avocado works with

0.2 s

to first sound with Kokoro

0.9 s

or less for a 42-second paragraph with Kokoro

0

accounts to create

Speeds measured with Kokoro on an M3 Max MacBook Pro, with the model already loaded.

Generate

Type it, hear it, keep the take you like

Choose a model and a voice, type or paste your script and press Generate. Playback starts while the rest is still being made, so you hear the first words almost straight away.

  • Direct the delivery. With models that support it, ask for "warm and unhurried" or "whispering", or drop emotion and sound tags such as laughs and sighs right into the script.
  • Pace and language. Slow a voice down, speed it up or pick the language, where the model offers it.
  • Takes, not overwrites. Every generation is kept as a take you can replay, compare and regenerate.

Create Voice

Clone, design or mix a voice of your own

Clone a voice from a recording you make in Avocado or clips you import. Only clone voices you have permission to use: Avocado asks you to confirm before it makes a clone.

Design a voice from a description, like "a calm narrator in her forties with a light Irish accent". Mix Kokoro's built-in voices by weight to find one that is in between.

How voice cloning works
Create Voice set up to clone a voice, listing what the model needs, including permission to clone the voice

Design

Describe a voice and hear it

Pick an age, accent, pitch, pacing and mood, or write the description yourself, then audition it before you save. Voice design works with Maya1, Qwen3-TTS Voice Design, VoxCPM 2 and Breeze TTS 2.

Voice design
Create Voice building a voice from age, gender, accent, pitch, timbre and pacing choices

Mix

Blend preset voices into a new one

Choose two to four of Kokoro's voices and set how much of each goes in. Audition the mix, save it, and it appears with your other voices, ready to use.

Mixing two Kokoro voices, Heart at 60 percent and Michael at 40 percent

Train

Teach it your own voice

Record or import between one and thirty minutes of your own speech and Avocado trains a voice from it, right on your Mac. Training works with a dozen model families, including Breeze TTS 2, Fish Audio S2 Pro, OpenAudio S1 Mini, Orpheus, Qwen3-TTS and VoxCPM.

Trained voices appear in your Voice Library with a Trained badge. As with cloning, only train on recordings you have permission to use.

Train a voice from your recordings

Dialogue

Conversations with a voice for every speaker

Write a script with turns for each speaker, cast a voice for each one and generate the whole conversation as one take. VibeVoice 1.5B handles up to four speakers; Dia, Sesame CSM and Fish Audio S2 Pro make dialogue too.

Multi-speaker dialogue
A two-speaker dialogue script with Ryan and Serena cast as speakers 1 and 2

Edit takes

Change the emotion after the fact

With Step Audio EditX on your Mac, open any take up to 30 seconds long and change its emotion, style or pace, add sounds like laughter or breathing, remove background noise or trim silence. Each edit is saved as a new take, so the original is always there.

Editing takes
Editing a take’s style to Serious with two passes; the original take is kept

Sound effects

Sound effects and ambience from a description

Describe a sound, like rain on a tin roof or a busy café, choose a length up to 60 seconds and generate it with MOSS Sound Effect. It lands in your project beside your voice takes.

Sound effects

Projects and export

Every take filed, tagged and ready to export

  • Projects keep takes together. New takes go to General until you choose a project, and deleted takes wait in Trash.
  • Tags on voices and takes make them easy to find, and takes inherit their voice's tags.
  • History and a command palette (⌘K) take you back to anything you made.
  • Export as WAV, AIFF or M4A, or drag a take out as a WAV.
The Projects page with an Audiobook project and takes tagged final, male and warm

Models

Works with the models you choose

Avocado works with 39 families of open voice models. Each one has its own strengths: some have dozens of ready voices, some clone, some take direction, some can be trained.

Also: Maya1, OuteTTS, Kitten TTS, Magpie TTS, VibeVoice Realtime, OpenAudio S1 Mini, Pocket TTS, IndexTTS, IndexTTS-2, MOSS TTS Nano, MOSS TTS Local, MOSS Sound Effect, VoxCPM 2, VoxCPM 1.5, Zonos v0.1, Zonos 2, Sesame CSM, Soprano, Dia, Echo TTS, Irodori TTS, Voxtral TTS, Chatterbox Multilingual, Chatterbox Turbo, Dramabox, LFM2.5 Audio, Step Audio EditX, StyleTTS 2 (LibriTTS), StyleTTS 2 (LJSpeech), XTTS v2, OpenVoice v2.

See every model

Privacy

Your voice stays on your Mac

  • Voices are generated on your Mac. Scripts, recordings, voices and takes stay in Avocado.
  • No account and nothing to sign in to.
  • Usage and diagnostics sharing is off until you choose, and never includes your scripts or audio.
  • Avocado goes online only to check for updates and to send feedback you write.
Private text to speech

Before you start

Bring your own voice models

Avocado works with popular open voice models that you add to your Mac. It doesn't include any models and doesn't download them for you, so there is a little setup: get the models you want from their makers, put them where Avocado looks, and they appear in the model picker.

Each model keeps its maker's licence. Some are free for any use, others are for research or non-commercial use only, so check the maker's page before you publish. Avocado checks each model on your Mac and shows you which are ready.

Questions

Is Avocado free?

Yes. Avocado will be free for Apple silicon Macs, with no account and no subscription. It is coming soon.

Which Macs will it run on?

Macs with Apple silicon (M1 or later) running macOS 15 or later. Bigger voice models need more memory, so a Mac with 16 GB or more gives you the most choice.

Does Avocado come with voices?

Avocado works with popular open voice models that you add to your Mac yourself. It does not include any models and does not download them for you. Once a model is on your Mac, Avocado finds it, shows what it can do and uses its voices.

Can I clone my own voice?

Yes, with models that support it. Record or import a few seconds of clear speech, add what was said, and save it as a voice. Only clone voices you have permission to use: Avocado asks you to confirm that before it makes a clone.

Can I use what I make commercially?

That depends on the model. Each voice model has its own licence from the people who made it, and some are for research or non-commercial use only. Check the model maker’s page before you publish anything.

Are my scripts and recordings uploaded?

No. Voices are generated on your Mac, and your scripts, recordings and takes stay in Avocado. Avocado goes online only to check for updates and to send feedback you write yourself.

What can I export?

Any take, as WAV, AIFF or M4A. You can also drag a take straight out of Avocado into another app as a WAV file.

Avocado is almost here

Free for Apple silicon Macs running macOS 15 or later. Coming soon.

Coming soon for Mac