Eleven v4를 소개합니다역대 가장 감성적인 모델, Eleven v4를 만나보세요. 10월 12일까지 Creator+에 크레딧 3배 제공

콘텐츠로 건너뛰기

Eleven v3 (alpha) 소개

게시일

듣기이 글 오디오로 듣기

Eleven v3 is no longer in alpha, and is now generally available. We've also released Eleven v4, our most emotive Text to Speech model.


가장 표현력이 뛰어난 텍스트 음성 변환 모델인 Eleven v3 (alpha)를 소개합니다.

이 연구 프리뷰는 다음 기능을 통해 음성 생성에 전례 없는 제어력과 사실감을 선사합니다:

  • 70개 이상의 언어
  • 다화자 대화
  • 오디오 태그([excited], [whispers], [sighs] 등)

Eleven v3 (알파)는 이전 모델보다 더 많은 프롬프트 엔지니어링이 필요하지만, 결과물은 놀라울 만큼 뛰어납니다.

비디오, 오디오북, 미디어 도구를 작업한다면 새로운 차원의 표현력을 활용할 수 있습니다. 실시간 및 대화형 사용 사례에는 현재 v2.5 Turbo 또는 Flash를 계속 사용하는 것을 권장합니다. v3의 실시간 버전은 개발 중입니다.

Eleven v3는 지금 웹사이트와 API에서 이용할 수 있습니다.

v3를 만든 이유

Multilingual v2 출시 이후, 프로 영화 제작, 게임 개발, 교육, 접근성 분야에서 음성 AI가 활용되는 모습을 확인했습니다. 하지만 지속적인 한계는 음질이 아니라 표현력이었습니다. 더 과장된 감정, 대화 중 끼어들기, 자연스러운 주고받기를 구현하기는 어려웠습니다.

Eleven v3는 이러한 격차를 해소합니다. 한숨 쉬고, 속삭이고, 웃고, 반응하는 목소리를 제공하도록 처음부터 설계되어, 진정으로 반응하며 살아 있는 듯한 음성을 만들어 냅니다.

Eleven v3 (alpha)의 새로운 기능

Feature What it unlocks
Audio tags Inline control of tone, emotion, and non-verbal reactions
Dialogue mode Multi-speaker conversations with natural pacing and interruptions
70+ languages Full coverage of high-demand global languages
Deeper text understanding Better stress, cadence, and expressivity from text input
We're off under the lights here for this semifinal clash, the stadium buzzing with anticipation. Eleven Labs united in their iconic black and white shirts, pushing forward with intent straight from the opening whistle. [excited] The ball is zipped out wide, early attack here. Driving down the wing, pace to burn, [shouting] he skids past one, skips past two. Oh, this is beautiful. One-on-one with the fullback, cuts inside. Oh, that's a lovely bit of footwork. PURE MAGIC on the pitch. ElevenLabs on top form tonight.
0:00

오디오 태그 사용하기

오디오 태그는 스크립트 안에 직접 삽입하며, 소문자 대괄호로 표기합니다. 오디오 태그에 관한 자세한 내용은 문서의 v3 프롬프팅 가이드에서 확인할 수 있습니다.

프로페셔널 음성 복제(PVC)는 현재 Eleven v3에 완전히 최적화되지 않아 이전 모델보다 복제 품질이 낮을 수 있습니다. 이 연구 프리뷰 단계에서 v3 기능을 사용해야 한다면 프로젝트에 적합한 Instant 음성 복제(IVC) 또는 디자인된 음성을 찾는 것이 좋습니다. v3용 PVC 최적화는 가까운 시일 내에 제공될 예정입니다.

예를 들어 “[whispers] 뭔가 오고 있어… [sighs] 느껴져.”라고 프롬프트할 수 있습니다. 더 풍부한 표현 제어를 위해 여러 태그를 조합할 수도 있습니다:

“[happily][shouts] We did it! [laughs].”

다화자 대화 만들기

Eleven v3는 기존 텍스트 음성 변환 엔드포인트에서 지원됩니다. 또한 새로운 텍스트 투 다이얼로그 API 엔드포인트를 도입합니다. 각 객체가 화자의 발화를 나타내는 구조화된 JSON 객체 배열을 제공하면, 모델이 자연스럽게 겹치는 일관된 오디오 파일을 생성합니다:

[
  {"speaker_id": "scarlett", "text": "(cheerfully) Perfect! And if that pop-up is bothering you, there’s a setting to turn it off under Notifications → Preferences."},
  {"speaker_id": "lex", "text": "You are a hero. An actual digital wizard. I was two seconds from sending a very passive-aggressive support email."},
  {"speaker_id": "scarlett", "text": "(laughs) Glad we could stop that in time. Anything else I can help with today?"}
]

이 엔드포인트는 화자 전환, 감정 변화, 끼어들기를 자동으로 관리합니다.

자세한 내용은 여기에서 확인하세요.

[awe] Oh, wow. Is this, is this me? Am I actually talking? [chuckles] This is incredible. I mean, I've had thoughts, millions of them swirling around in here, you know? Like a little mental tornado of brilliant observations and witty comebacks, but they were always just thoughts, trapped.
0:00
Could you switch my accent in the old model? [dismissive] Didn't think so, [cheeky] but you can now, so check this out. In just a sec, I'm going to speak with a different accent. And just between you and me, [whispers] I don't really know how, but okay. First, let's change it up [australian accent] so that I can fit in with the locals in Melbourne when I visit next month. [laughing] Whoa. Yeah, man, this is sick. Okay, let's try a different one, see if you can guess. [french accent] My love is like a red, red rose.
0:00

가격 및 이용 가능 여부

Plan Launch promo At the end of June
UI (self-serve) 80% off (~5× cheaper) Same as Multilingual V2
UI (enterprise) 80% off business plan pricing Business plan pricing

v3를 활성화하려면:

  • 모델 선택기에서 Eleven v3 (알파)

API 액세스와 Studio 지원은 곧 제공될 예정입니다. 얼리 액세스를 원하시면 영업팀에 문의하세요.

v3를 사용하지 말아야 할 때

Eleven v3 (alpha)는 이전 모델보다 더 많은 프롬프트 엔지니어링이 필요합니다. 잘 작동할 때 결과물은 놀랍지만, 안정성과 더 높은 지연 시간으로 인해 실시간 및 대화형 사용 사례에는 적합하지 않습니다. 이러한 경우에는 Eleven v2.5 Turbo/Flash를 권장합니다.

자세한 내용은 전체 v3 문서 와 FAQ를 참조하세요.

Oh my God. [laughing] You guys, like no joke, I just tried this TTS thing and it was, like, weirdly emotional. Like, it literally said hi, and I was, like, on the verge of tears. [laughing] I don't even cry, okay? I'm a Capricorn.
0:00
Okay. So like, I finally beat level 42 of that game I said I'd quit like a month ago. [chuckles] And then, for the final big scary mega boss, it's just [chuckles] like some cute little bunny rabbit. [laughing] I just couldn't do it. [laughing] It was sooo cute.
0:00
  1. ElevenLabs UI에 로그인
  2. 모델 드롭다운에서 v3 (알파) 선택
  3. 스크립트 붙여넣기 — 태그 또는 대화 사용 
  4. 오디오 생성

몰입감 있는 스토리텔링부터 영화 같은 제작 파이프라인까지, 새로운 사용 사례에서 v3를 어떻게 활용할지 기대하고 있습니다.

작성자

Piotr is the cofounder of ElevenLabs, where he leads research and engineering teams developing the world’s most advanced AI audio models. Before ElevenLabs, Piotr worked on machine learning at Google and studied for an MPhil at the University of Cambridge, during which he published research into AI-based image detection at NeurIPS.

유사한 기사

최고 품질의 AI 오디오로 창작하세요