ElevenLabs’ new v4 speech model supports more expression control and 90 languages
ElevenLabs launched its v4 and v4 Turbo speech models, enhancing expression control and supporting over 90 languages. The new architecture allows voice cloning with just 10 seconds of audio. This update also reduces latency for voice agents, making conversations more fluid. The company has seen its annualized revenue run rate climb to over $600 million.
ElevenLabs' new v4 model is less about language expansion and more about a rapid revenue ramp. The company’s annualized revenue run rate has surged from $330 million to over $600 million this year. This growth, alongside a recent $500 million raise valuing it at $11 billion, points to a clear market pull for advanced voice AI.
The focus on Japanese, Brazilian Portuguese, Mandarin, and Cantonese quality jumps directly impacts key Asian markets. This improved capability could accelerate adoption for voice agents in customer service and content creation across China, Japan, and other regions. Asian companies will find it easier to deploy sophisticated, localized voice AI solutions.
The thing to watch is the rumored $22 billion valuation in a follow-up round. Such a valuation would cement ElevenLabs as a dominant player, pushing competitors like Google and OpenAI to innovate faster. It also sets a high bar for other voice AI startups in Asia seeking similar investor confidence.
Share this article
Related reading
6 stories
How Can Banks Launch New Products Without Replacing Their Core?

Wise Rolls Out Overseas QR Payments, Customisable eSIM Plans

NYC Council Speaker Julie Menin warns of AI's 'existential' risks

What Are OpenAI Dots And Do You Need Them?

Google to roll out new Map tool in Gemini app to help with location-based prompts

