跳到内容

Text to Speech

Text to Speech with high quality, human-like AI voices

Log in with Google

Trusted by 1M+ users • Free to start

旁白

富有表现力的音色,让有声书和播客更生动。

广告

有说服力的音色,助力转化和品牌记忆

角色

适合动画或游戏的生动有趣音色。

旁白

富有表现力的音色,让有声书和播客更生动。

对话

适用于非正式场景的自然音色。

社交媒体

适合短内容的潮流吸睛音色。

Emotionally & contextually aware AI voices for Text to Speech

Our voice AI responds to emotional cues in text and adapts its delivery to suit both the immediate content and the wider context. This lets our AI voices achieve high emotional range and avoid making logical errors when your content is read aloud.

Control the emotion, delivery and direction

Create controllable, expressive speech layered with emotion, audio events, and immersive soundscapes.

Access a library of 10,000+ human-like voices

Explore an ever-growing collection of expressive, lifelike voices for any use case - from narration to character creation.

Dialogue support

Create audio conversations where speakers share context and emotion.

Clone or design a voice

Instantly replicate your own voice or craft unique AI Voices with full control.

Multilingual speech

Bring stories to life in over 70 languages, all with native-level emotion and clarity.

Built for a wide range of use cases, from AI Agents to audiobooks or voiceovers

Conversational Agents

Use AI text to speech to create natural, human-like voices for chatbots and virtual assistants, improving user interaction with realistic responses.
TTS used for Conversational Agents

Gaming

Generate voiceovers for video game characters using the text to speech API, with context-aware and emotionally accurate voices that match in-game scenarios.
Text to Speech for narration in video games

Audiobooks

Convert written text into natural-sounding AI voices for audiobooks, allowing you to produce content quickly in multiple languages.
Use TTS to generate audiobooks

Video voiceovers

Produce high-quality voiceovers for videos, TV shows, and animations using AI text to voice, eliminating the need for human voice actors and speeding up production.
Add voiceovers to videos with TTS

Podcasts

Use AI text to speech for creating podcasts with consistent, professional-sounding narration, reducing the time spent on manual recording.
Generate podcasts in Studio

Accessibility

Integrate text to speech into websites and apps to provide audio versions of content, helping users with visual impairments or reading difficulties access information more easily.
Use text to speech for screen readers

Millions of words generated every minute

Generate speech in over 70 languages and wide range of accents

  • English
  • Afrikaans
  • Arabic
  • Armenian
  • Assamese
  • Azerbaijani
  • Belarusian
  • Bengali
  • Bosnian
  • Bulgarian
  • Catalan
  • Cebuano
  • Chichewa
  • Chinese
  • Croatian
  • Czech
  • Danish
  • Dutch
  • Estonian
  • Filipino
  • Finnish
  • French
  • Galician
  • Georgian
  • German
  • Greek
  • Gujarati
  • Hausa
  • Hebrew
  • Hindi
  • Hungarian
  • Icelandic
  • Igbo
  • Indonesian
  • Irish
  • Italian
  • Japanese
  • Javanese
  • Kannada
  • Kazakh
  • Kirghiz
  • Korean
  • Latvian
  • Lingala
  • Lithuanian
  • Luxembourgish
  • Macedonian
  • Malay
  • Malayalam
  • Mandarin Chinese
  • Marathi
  • Nepali
  • Norwegian
  • Pashto
  • Persian
  • Polish
  • Portuguese
  • Punjabi
  • Romanian
  • Russian
  • Serbian
  • Sindhi
  • Slovak
  • Slovenian
  • Somali
  • Spanish
  • Swahili
  • Swedish
  • Tamil
  • Telugu
  • Thai
  • Turkish
  • Ukrainian
  • Urdu
  • Vietnamese
  • Welsh
  • African
  • American
  • Argentine
  • Australian
  • British
  • Californian
  • Canadian
  • Cockney
  • Country
  • Czech Moravian
  • Filipino
  • French Swiss
  • German
  • German Bavarian
  • Indian
  • Irish
  • Italian
  • Latin American
  • Latino
  • Mexican
  • New York
  • Pakistani
  • Portuguese
  • Russian
  • Scandinavian
  • Scottish
  • Southern USA
  • Spanish
  • Spanish Castilian
  • Texas
  • Turkish Istanbul

Built on the most powerful Text to Speech models

Eleven v3

Our most advanced, expressive model with audio tags for precise emotional control. Best for storytelling, gaming and media production in 70+ languages.

  • Dramatic delivery and performance
  • 70+ languages supported
  • 5,000 character limit
  • Multi-speaker dialogue

Multilingual v2

Our most lifelike, emotionally rich text to speech model supporting 29 languages. Best for voiceovers, audiobooks, post-production and content creation.

  • Natural-sounding output
  • 29 languages supported
  • 10,000 character limit
  • Designed for long-form generations

Flash v2.5

Our high quality, low latency TTS model in 32 languages. Best for developer use cases where speed matters and you need non-English languages

  • Ultra-low latency (~75ms*)
  • 32 languages supported
  • 40,000 character limit
  • Faster model, 50% lower price per character

Turbo v2.5

High quality, low-latency model with a good balance of quality and speed

  • High quality voice generation
  • 32 languages supported
  • 40,000 character limit
  • Low latency (~250ms-300ms†), 50% lower price per character

Trusted by the World’s Leading Creators & their communities

Text to Speech Pricing

$0 每月

10k 积分 每月

免费
$0 每月
入门
$6 每月
创作者
$22 每月, 首月 $11
专业
$99 每月

10k 积分 每月

30k 积分 每月

121k 积分 每月

600k 积分 每月

文本转语音(界面)

包含分钟数

额外分钟数

音频质量

~10

~$0.36

128 kbps,44.1kHz

~30

~$0.20

128 kbps,44.1kHz

~121

~$0.18

128 kbps,44.1kHz

~600

~$0.17

128 & 192 kbps(Studio & API),44.1kHz

价格不含税费。 查看全部定价方案。

Enterprise-grade security and infrastructure at scale

Foreground

Available on the web, mobile and via APIs or SDKs

ElevenLabs Studio

The best AI audio models in one powerful editor.

Use TTS in the ElevenLabs Studio

ElevenLabs Mobile App

Generate expressive audio in seconds using our iOS and Android apps.

ElevenLabs Mobile app

Text to Speech APIs and SDKs

Integrate ElevenLabs Text to Speech (TTS) into your product via APIs or SDKs.

TTS API

Showcasing the global impact of AI audio research

Explore our AI voices for Text to Speech

常见问题

Latest updates

用高质量 AI 音频创作