• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
AI
Announcement

ElevenLabs introduces Eleven v4 and v4 Turbo voice models

ElevenLabs says v4 lets creators direct speech delivery with inline tags, while v4 Turbo is built for real-time use with a median inference latency of about 100 milliseconds.

Robert ScobleRS
Justine MooreJM
ElevenLabs DevelopersED
36 Sources, 12d ago, first seen 12d ago

TLDR

ElevenLabs introduced Eleven v4 and v4 Turbo, saying the models support 90-plus languages. The company says creators can use inline tags to direct v4’s delivery, emotion and pacing, while v4 Turbo is designed for real-time use. Artificial Analysis ranks Eleven v4 first on its Provider Voice leaderboard and Pronunciation Robustness benchmark, and second on Controlled Voice.

Combined views

152.9K

36 Sources, first seen 12d ago

1.4K likes45 comments804 saves75 reposts

Combined views

152.9K

36 Sources, first seen 12d ago

1.4K likes45 comments804 saves75 reposts

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Featured Source

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Related

ElevenLabs introduces ElevenAgents Architect for building and improving AI agents

The company says proposed changes stay in versioned drafts for review, simulation and rollback before launch.

Introducing ElevenAgents Architect
ElevenAgents joins the OpenAI Marketplace

ElevenLabs says OpenAI enterprise customers will soon be able to give their companies a voice in 70-plus languages and put ElevenAgents on their existing OpenAI commitment.

Eleven v4 Turbo is now available in ElevenAgents for live conversations

ElevenLabs says ElevenAgents can pick up on how a caller is feeling and adjust its delivery, with more expression and consistency on long calls.

36 Sources

ElevenLabs@ElevenLabsIntroducing Eleven v4 and Eleven v4 Turbo, our fastest and most emotive voice models yet. Ranked #1 by Artificial Analysis.12d
Mati Staniszewski@matiEleven v4 & v4 Turbo Text to Speech models are here! - v4: #1 quality TTS, ~100 languages, new level of expressivity & control - v4 Turbo: fastest TTS, ~100ms TTFB, optimized for real-time conversations Live now via API & all our products - at $22/$11 per 1M chars for 2 week!12d
ElevenLabs Developers@ElevenLabsDevsIntroducing Eleven v4 and Eleven v4 Turbo, our fastest and most emotive voice models yet. Ranked #1 by Artificial Analysis, Eleven v4 delivers the most expressive results for developers building creative experiences, while Eleven v4 Turbo is optimized for real-time voice agents.12d
Piotr Dabkowski@dabkowski_piotrA couple of things that make Eleven v4 special: - Extremely low latency (100 ms p50 TTFB for Turbo) - By far the best voice similarity - Outstanding quality (#1 on Artificial Analysis and on our own benchmarks) It achieves all of this while being extremely reliable. Eleven v4 uses our new post-training stack and a surprising recent research discovery that we hope will greatly accelerate our future model development. By the end of the year, we expect delivery of new compute that will increase our training capacity almost 10x. Very excited for what's ahead! Congratulations to the team!12d
Artificial Analysis@ArtificialAnlysElevenLabs’ Eleven v4 takes #1 on the Artificial Analysis Provider Voice TTS Arena Leaderboard and Pronunciation Robustness benchmark, and #2 on Controlled Voice, surpassing Cartesia’s Sonic 3.6 and Google’s Gemini 3.8 Flash TTS on Provider Voice Eleven v4 is the latest Text to Speech model from @ElevenLabs, with support for 90+ languages, up from 70+ for Eleven v3. Key takeaways: ➤ Provider Voice: Eleven v4 takes #1 with an Elo of 1,319 (+19/-19) across 1,674 appearances, ahead of Cartesia’s Sonic 3.6 at 1,276 and Google’s Gemini 3.8 Flash TTS at 1,267. It ranks #1 across all four categories: Customer Service, Assistants, Knowledge Sharing and Entertainment. ➤ Controlled Voice (every model uses the same custom voice for comparison): Eleven v4 ranks #2 with an Elo of 1,157 (+16/-16) across 1,483 appearances, just behind Alibaba’s Qwen-Audio-3.1-TTS-Plus at 1,178 and well ahead of Eleven v3 at 1,073. ➤ Pronunciation Robustness: Eleven v4 scores 91.7%, the highest score we have measured, ahead of Gemini 3.8 Flash TTS at 89.5% and Gemini 3.1 Flash TTS at 88.2%, and up from Eleven v3 at 85.6%. ➤ Cost: Eleven v4 costs $80/1M characters, compared to $49/1M characters for Sonic 3.6 and $16.49/1M characters for Gemini 3.8 Flash TTS. ➤ Speed: Eleven v4 processes 73.4 characters per second of generation time, compared to 42.5 characters per second for Eleven v3. See more details and listen to samples below 🧵12d
Robert Scoble@ScobleizerRT @ElevenLabsDevs: Introducing Eleven v4 and Eleven v4 Turbo, our fastest and most emotive voice models yet. Ranked #1 by Artificial Anal…12d
Justine Moore@venturetwinsTruly blown away by the results from Eleven v4. This model has a new architecture that unlocks more realistic and controllable speech. You can now direct the performance of the character AND the soundscape around them. And the voice effects (like "cheap microphone") are 👌12d
Nev Flynn@NevFlynnEleven v4 is our most emotive model. We built a demo to show you its emotional range, where [audio tags] shape the feeling of every prompt.12d
🚨 AI News | TestingCatalog@testingcatalogElevenLabs launched Eleven v4 and Eleven v4 Turbo, with Eleven v4 taking a first spot on the Artificial Analysis Provider Voice TTS Arena. > "Eleven v4 is built on a new architecture to generate speech with the intended tone, pacing, emotion, and character." A new TTSOTA 👀12d
AI Tensibility@AITensibilityElevenLabs เปิดตัว Eleven v4 และ v4 Turbo เสียง AI สมจริงขึ้น โคลนเสียงได้ใน 10 วินาที ElevenLabs เปิดตัวโมเดลเสียง AI รุ่นใหม่ Eleven v4 และ Eleven v4 Turbo ซึ่งบริษัทระบุว่าเป็นโมเดลเสียงที่เร็วและถ่ายทอดอารมณ์ได้ดีที่สุดของตนในปัจจุบัน โดยโมเดลได้รับการจัดอันดับเป็นอันดับ 1 จาก Artificial Analysis หัวใจสำคัญของ Eleven v4 คือสถาปัตยกรรมใหม่ที่ออกแบบมาเพื่อสร้างเสียงพูดให้สอดคล้องกับสิ่งที่ผู้สร้างต้องการมากขึ้น ทั้งน้ำเสียง จังหวะการพูด อารมณ์ และบุคลิกของตัวละคร ทำให้เสียงที่สร้างขึ้นมีลักษณะใกล้เคียงกับการแสดงของมนุษย์มากกว่า Text-to-Speech แบบเดิมที่มักให้เสียงค่อนข้างราบเรียบ ขณะที่ Eleven v4 Turbo ถูกพัฒนาสำหรับงานแบบ Real-time โดยเฉพาะ มีค่าความหน่วงในการประมวลผลระดับประมาณ 100 มิลลิวินาที ทำให้เหมาะกับระบบที่ต้องตอบสนองทันที เช่น AI Voice Agent งานบริการลูกค้า ฝ่ายขาย ระบบนัดหมาย หรือผู้ช่วยเสียงที่ต้องสนทนากับผู้ใช้อย่างเป็นธรรมชาติ อีกหนึ่งการพัฒนาสำคัญคือความสามารถด้าน Voice Cloning ที่มีความสมจริงและสม่ำเสมอมากขึ้น โดย Instant Voice Clone สามารถเรียนรู้ลักษณะเสียงจากตัวอย่างเสียงเพียงประมาณ 10 วินาที ขณะที่ Professional Voice Clone ถูกออกแบบสำหรับงานที่ต้องการคุณภาพและความเหมือนของเสียงในระดับสูงสุด สำหรับงานสร้างคอนเทนต์แบบยาว Eleven v4 ยังถูกออกแบบให้รักษาบุคลิก น้ำเสียง และลักษณะการพูดให้ต่อเนื่องสม่ำเสมอ ช่วยลดปัญหาเสียงเปลี่ยนโทนหรือบุคลิกระหว่างการสร้างเสียงยาว ๆ เช่น หนังสือเสียง พอดแคสต์ วิดีโอ หรือบทสนทนาหลายฉาก นอกจากนี้ ผู้สร้างสามารถกำกับการแสดงของเสียงผ่าน Inline Tags ได้โดยตรง เช่น กำหนดอารมณ์ จังหวะ ความเร็ว ปฏิกิริยา เสียงประกอบ รูปแบบการพูด หรือแม้แต่เขียนคำอธิบายว่าต้องการให้ประโยคนั้นถูกถ่ายทอดอย่างไร จากนั้น Eleven v4 จะนำคำสั่งเหล่านี้ไปสร้างการแสดงเสียงให้สอดคล้องกับบริบทของฉาก แหล่งที่มา: ข้อมูลเพิ่มเติม: https://elevenlabs.io/v4 ทดสอบใช้งาน: https://elevenlabs.io/app/speech-synthesis/text-to-speech11d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    ElevenLabsEleven v4Eleven v4 Turbo
    Artificial Analysis

    36 Sources

    ElevenLabs@ElevenLabsIntroducing Eleven v4 and Eleven v4 Turbo, our fastest and most emotive voice models yet. Ranked #1 by Artificial Analysis.12d
    Mati Staniszewski@matiEleven v4 & v4 Turbo Text to Speech models are here! - v4: #1 quality TTS, ~100 languages, new level of expressivity & control - v4 Turbo: fastest TTS, ~100ms TTFB, optimized for real-time conversations Live now via API & all our products - at $22/$11 per 1M chars for 2 week!12d
    ElevenLabs Developers@ElevenLabsDevsIntroducing Eleven v4 and Eleven v4 Turbo, our fastest and most emotive voice models yet. Ranked #1 by Artificial Analysis, Eleven v4 delivers the most expressive results for developers building creative experiences, while Eleven v4 Turbo is optimized for real-time voice agents.12d
    Piotr Dabkowski@dabkowski_piotrA couple of things that make Eleven v4 special: - Extremely low latency (100 ms p50 TTFB for Turbo) - By far the best voice similarity - Outstanding quality (#1 on Artificial Analysis and on our own benchmarks) It achieves all of this while being extremely reliable. Eleven v4 uses our new post-training stack and a surprising recent research discovery that we hope will greatly accelerate our future model development. By the end of the year, we expect delivery of new compute that will increase our training capacity almost 10x. Very excited for what's ahead! Congratulations to the team!12d
    Artificial Analysis@ArtificialAnlysElevenLabs’ Eleven v4 takes #1 on the Artificial Analysis Provider Voice TTS Arena Leaderboard and Pronunciation Robustness benchmark, and #2 on Controlled Voice, surpassing Cartesia’s Sonic 3.6 and Google’s Gemini 3.8 Flash TTS on Provider Voice Eleven v4 is the latest Text to Speech model from @ElevenLabs, with support for 90+ languages, up from 70+ for Eleven v3. Key takeaways: ➤ Provider Voice: Eleven v4 takes #1 with an Elo of 1,319 (+19/-19) across 1,674 appearances, ahead of Cartesia’s Sonic 3.6 at 1,276 and Google’s Gemini 3.8 Flash TTS at 1,267. It ranks #1 across all four categories: Customer Service, Assistants, Knowledge Sharing and Entertainment. ➤ Controlled Voice (every model uses the same custom voice for comparison): Eleven v4 ranks #2 with an Elo of 1,157 (+16/-16) across 1,483 appearances, just behind Alibaba’s Qwen-Audio-3.1-TTS-Plus at 1,178 and well ahead of Eleven v3 at 1,073. ➤ Pronunciation Robustness: Eleven v4 scores 91.7%, the highest score we have measured, ahead of Gemini 3.8 Flash TTS at 89.5% and Gemini 3.1 Flash TTS at 88.2%, and up from Eleven v3 at 85.6%. ➤ Cost: Eleven v4 costs $80/1M characters, compared to $49/1M characters for Sonic 3.6 and $16.49/1M characters for Gemini 3.8 Flash TTS. ➤ Speed: Eleven v4 processes 73.4 characters per second of generation time, compared to 42.5 characters per second for Eleven v3. See more details and listen to samples below 🧵12d
    Robert Scoble@ScobleizerRT @ElevenLabsDevs: Introducing Eleven v4 and Eleven v4 Turbo, our fastest and most emotive voice models yet. Ranked #1 by Artificial Anal…12d
    Justine Moore@venturetwinsTruly blown away by the results from Eleven v4. This model has a new architecture that unlocks more realistic and controllable speech. You can now direct the performance of the character AND the soundscape around them. And the voice effects (like "cheap microphone") are 👌12d
    Nev Flynn@NevFlynnEleven v4 is our most emotive model. We built a demo to show you its emotional range, where [audio tags] shape the feeling of every prompt.12d
    🚨 AI News | TestingCatalog@testingcatalogElevenLabs launched Eleven v4 and Eleven v4 Turbo, with Eleven v4 taking a first spot on the Artificial Analysis Provider Voice TTS Arena. > "Eleven v4 is built on a new architecture to generate speech with the intended tone, pacing, emotion, and character." A new TTSOTA 👀12d
    AI Tensibility@AITensibilityElevenLabs เปิดตัว Eleven v4 และ v4 Turbo เสียง AI สมจริงขึ้น โคลนเสียงได้ใน 10 วินาที ElevenLabs เปิดตัวโมเดลเสียง AI รุ่นใหม่ Eleven v4 และ Eleven v4 Turbo ซึ่งบริษัทระบุว่าเป็นโมเดลเสียงที่เร็วและถ่ายทอดอารมณ์ได้ดีที่สุดของตนในปัจจุบัน โดยโมเดลได้รับการจัดอันดับเป็นอันดับ 1 จาก Artificial Analysis หัวใจสำคัญของ Eleven v4 คือสถาปัตยกรรมใหม่ที่ออกแบบมาเพื่อสร้างเสียงพูดให้สอดคล้องกับสิ่งที่ผู้สร้างต้องการมากขึ้น ทั้งน้ำเสียง จังหวะการพูด อารมณ์ และบุคลิกของตัวละคร ทำให้เสียงที่สร้างขึ้นมีลักษณะใกล้เคียงกับการแสดงของมนุษย์มากกว่า Text-to-Speech แบบเดิมที่มักให้เสียงค่อนข้างราบเรียบ ขณะที่ Eleven v4 Turbo ถูกพัฒนาสำหรับงานแบบ Real-time โดยเฉพาะ มีค่าความหน่วงในการประมวลผลระดับประมาณ 100 มิลลิวินาที ทำให้เหมาะกับระบบที่ต้องตอบสนองทันที เช่น AI Voice Agent งานบริการลูกค้า ฝ่ายขาย ระบบนัดหมาย หรือผู้ช่วยเสียงที่ต้องสนทนากับผู้ใช้อย่างเป็นธรรมชาติ อีกหนึ่งการพัฒนาสำคัญคือความสามารถด้าน Voice Cloning ที่มีความสมจริงและสม่ำเสมอมากขึ้น โดย Instant Voice Clone สามารถเรียนรู้ลักษณะเสียงจากตัวอย่างเสียงเพียงประมาณ 10 วินาที ขณะที่ Professional Voice Clone ถูกออกแบบสำหรับงานที่ต้องการคุณภาพและความเหมือนของเสียงในระดับสูงสุด สำหรับงานสร้างคอนเทนต์แบบยาว Eleven v4 ยังถูกออกแบบให้รักษาบุคลิก น้ำเสียง และลักษณะการพูดให้ต่อเนื่องสม่ำเสมอ ช่วยลดปัญหาเสียงเปลี่ยนโทนหรือบุคลิกระหว่างการสร้างเสียงยาว ๆ เช่น หนังสือเสียง พอดแคสต์ วิดีโอ หรือบทสนทนาหลายฉาก นอกจากนี้ ผู้สร้างสามารถกำกับการแสดงของเสียงผ่าน Inline Tags ได้โดยตรง เช่น กำหนดอารมณ์ จังหวะ ความเร็ว ปฏิกิริยา เสียงประกอบ รูปแบบการพูด หรือแม้แต่เขียนคำอธิบายว่าต้องการให้ประโยคนั้นถูกถ่ายทอดอย่างไร จากนั้น Eleven v4 จะนำคำสั่งเหล่านี้ไปสร้างการแสดงเสียงให้สอดคล้องกับบริบทของฉาก แหล่งที่มา: ข้อมูลเพิ่มเติม: https://elevenlabs.io/v4 ทดสอบใช้งาน: https://elevenlabs.io/app/speech-synthesis/text-to-speech11d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet