Subscribe Sign in

ElevenLabs’ new v4 speech model supports more expression control and 90 languages

TechCrunch
1 min read Rewritten in plain language

AppsVoice AI

Show what we removed
  • ElevenLabs launched two new speech models on Monday, called ElevenLabs v4 and v4 Turbo, offering more expression control, lower latency for voice agents, and support for more than 90 languages.
  • For the v4 generation of models, ElevenLabs is adopting a new architecture that allows for better control and faster cloning.
  • The previous version supported 70 languages, and ElevenLabs has worked to get that number up to 90 languages with the new version.
  • ElevenLabs raised $500 million earlier this year from Sequoia earlier this year, in a round led by Sequoia that valued the company at $11 billion.

4 sentences from our version of the report, chosen to cover it. Nothing here is written; every line is in the article below. How

Headline check

The one thing this headline claims is in the report.

Figures, names and quoted words in the headline, looked for in the report itself — not in the summary above. One claim in this headline could be checked, so this is a narrow pass and not a thorough one. How this is checked

Manns' superior seeds (15767599714) library picture
Not from this story. A library photograph of stock market trading screens, used to illustrate it. Manns' superior seeds (15767599714) Henry G. Gilbert Nursery and Seed Trade Catalog Collection.; J. Manns & Co. / Wikimedia Commons, CC BY

ElevenLabs launched two new speech models on Monday, called ElevenLabs v4 and v4 Turbo, offering more expression control, lower latency for voice agents, and support for more than 90 languages.

The company released its v3 model last year, and teased the model at an event in Warsaw earlier this year. For the v4 generation of models, ElevenLabs is adopting a new architecture that allows for better control and faster cloning. The company said that with v4, users will be able to clone a voice with just 10 seconds of audio.

ElevenLabs introduced inline tags to define expression with v3, and is expanding those tags in v4, letting users stack multiple tags and having the model follow the sequence.

The previous version supported 70 languages, and ElevenLabs has worked to get that number up to 90 languages with the new version. The startup said that it observed the biggest quality jump in Japanese, Brazilian Portuguese, Mandarin and Cantonese.

The company said that the new model is suited for voice agents, as the new version has lower latency to allow for more fluid conversation.

Competition in speech models has ramped up as startups like Cartesia, Deepgram, Fish Audio, Boson, and WellSaid Labs have created expressive speech models.

ElevenLabs raised $500 million earlier this year from Sequoia earlier this year, in a round led by Sequoia that valued the company at $11 billion.

Shortened to 1 minute of reading, this version reads 9.9 on the Niral Score.

You are reading our version, not theirs. This is TechCrunch's report shortened to its most important sentences, in plainer words, with verdicts and loaded words taken out. Plain description stays, and so do adjectives that carry a fact, such as "former" or "federal". The reporting, the facts and the quotations are theirs — quotations are never edited — and the indicators beside it measure this version. Hover or tap Adjectives to see every one left in the text.

How this outlet filed it, and how we rewrote it

No other newsroom we read has filed on this event, so there is nothing to compare it with yet.

Outlet Niral ScoreAdjectivesSourcingSentimentHappiness
TechCrunchas they published this story 11.5 12 26 0.1 55.4
Mundane Readneutralized from TechCrunch 11.5 12 26 0.1 55.4

Sign in to react.

Comments

Nothing here yet.

Sign in to comment.

Questions

Readers can ask a question about this story here. Questions and answers are for subscribers. Sign in to read them.

Comments are read before they appear where anything in them needs a person to look. Nothing posted here is ever deleted; a comment taken down keeps its text and the reason, so the decision can be looked at again. How this works