Introducing Multilingual Voices
We partnered with Baklavastory, our neighbor in the Mission, to show what it sounds like when a business's character and warmth stay consistent, no matter who calls or what language they speak.
It's easy to translate speech, but much harder to
We regret to inform our AI researcher Ronald that AI has replaced him.
Sorry Ronald.
Here's our team trying to tell his real voice from his cloned one. They were not good at it.
sonic-3.6 is out, with a large improvement over (the already #1) sonic-3.5 in just a few months! the research team's focus on fundamentals is accelerating progress at the frontier of architectures and audio
Sonic-3.6 is now generally available.
In January we made a bet: stop tuning the existing paradigm, rebuild from the architecture up.
How we topped our own best model in two months → cartesia.ai/blog/sonic-3.6…
the team continues cooking 👩🍳
this is now an unprecedented gap on both the super competitive Provider Voices leaderboard as well as the newer Controlled Voices leaderboard (a stronger benchmark that can't be benchmaxxed, requiring truly better algorithms)
Cartesia's TTS model is
Cartesia's Sonic 3.6 takes the #1 spot on both the Provider Voice and Controlled Voice Artificial Analysis Speech Arena leaderboards, surpassing Speechify AI's Simba 3.2 and Alibaba's Qwen-Audio-3.0-TTS-Plus, with Sonic 3.5 holding #2 on Controlled Voice
Sonic 3.6 is the latest