Voices
Your cloned voices, or a new voice for each seed
Tags
43, for emotion, style, pace, pauses and sounds
Languages
Many; strongest in English, Spanish, German, French, Italian, Portuguese and Russian
Voice cloning
From up to eight clips, 5 to 20 s in all
Long scripts
One voice held across the whole script
Memory
About 7.5 GB at the default setting
Made by
Boson AI

Higgs Audio v3 from Boson AI is built for expressive speech. It reads the mood from the text and from the tags you add.

Tags

The script’s tag button lists all 43 of Higgs’ tags in groups: emotion, style, pace and pitch, pauses and sounds such as laughter. Pick one and Avocado puts it in the right place: emotions at the start of the sentence, pauses and sounds where your cursor is.

Voices

There are no preset voices. Without a cloned voice, Higgs makes up a new voice for each seed, and the same seed always gives the same voice. Clone a voice from up to eight clips to use one consistently; any voice you’ve cloned for another model works too. Only clone voices you have permission to use.

Long scripts

Send a whole script at once. Higgs keeps one voice across it, and if a part stalls Avocado re-tries it quietly rather than failing the take.

Playback

Takes play as they’re made, so you start hearing the result before the whole script is finished.

Licence

Higgs Audio v3’s licence is set by Boson AI. Read it on the Higgs Audio v3 model page before you publish anything you make.

Questions

Why does the voice change when I generate again?

Without a cloned voice, Higgs invents a voice for each seed. Pin the seed to keep a voice you like, or clone one to use every time.

Does it take direction?

Not as a written instruction. Use the tags instead; Avocado places each one where it belongs in the sentence.

Avocado is almost here

Free for Apple silicon Macs running macOS 15 or later. Coming soon.

Coming soon for Mac