推出 Eleven v4认识 Eleven v4:迄今情感表现最丰富的模型。 Creator+ 套餐含 3 倍点数优惠,截止至 10 月 12 日

跳至内容

Eleven v3 Audio Tags:为语音中的角色表演提供指导

发布时间
最近更新

收听收听本文

Audio Tags 是 Eleven v3(alpha)这款全新研究预览版 文本转语音 模型提供的强大工具。它不仅能精准控制语调和节奏,还能指导角色与声音表演。 

借助 [pirate voice]、[French accent] 或 [sarcastically] 等标签,语音不再只是旁白,更能成为讲故事的工具。结合出色的角色语音克隆,不仅能捕捉声音,更能呈现完整表演。

这些标签可让你在一句话中切换声音身份、模仿口音,或呈现反派、旁白、配角等原型角色,无需修改原有脚本或切换音色。

AI 语音中的角色表演是什么?

角色表演就是进入某个角色的能力。无论是为张扬的反派、粗犷的海船船长,还是墨尔本当地店主配音,新的 Audio Tags 都能让你引导表达方式,贴合想呈现的人物形象。

只需一句简单的方括号提示,就能设定场景:“[pirate voice] 啊哈,辽阔的大海。闻到了吗,伙计们?那是自由的气息……还有一丝叛变的味道。”

模型不只是念出文字,还会以角色身份来表演。

从口音到角色原型

Arr, the open ocean. Smell that, lads? That's the scent of freedom and just a hint of mutiny. [laughs wickedly] Now grab your cutlasses, stow your fear. Tonight, we dine like kings or we sink like legends. [evil laugh]
0:00

声音表演不只关乎音量或情绪,也关乎是谁在说话。借助 Eleven v3,你可以随时提示特定口音、方言和说话风格。例如:

[American accent] 在旧模型里,你能切换我的口音吗?[dismissive] 我想不能吧。[Australian accent] 但现在可以了——看看这个,伙计![French accent] 我的爱……就像一朵鲜红的玫瑰。

这种流畅的身份切换非常适合动画、游戏、互动小说,以及任何说话者个性重要的场景。

常用角色表演标签

以角色为中心的标签可塑造声音身份与表现力:

  • 口音和方言: [British accent]、[Australian accent]、[Southern US accent]
  • 角色原型和身份: [pirate voice]、[evil scientist voice]、[childlike tone]
  • 说话风格: [dramatic]、[sarcastically]、[matter-of-fact]、[whiny]
  • 类型提示: [fantasy narrator]、[sci-fi AI voice]、[classic film noir]

叠加标签有助于让角色鲜活起来:“[dramatic][French accent] 你不明白……这从来不是为了复仇,而是命运。”

从旁白到群像配音

在多角色脚本中,Audio Tags 可让你轻松切换不同声音。只需在对话中切换角色表演,就能加入紧张感、幽默或惊喜,无需额外编辑。

[excited] Yo, Jessica. Oh my goodness, have you tried the new ElevenLabs v3?
[chuckles] Hey, Dr. Von Fusion. Yeah, I just got it. The clarity is amazing. Like, I can actually do whispers now, [whispers] like this.
[sarcastically] Ooh, well, look at you, miss fancy pants. Hey, check this out. I can do full Shakespeare now. [dramatically] To be or not to be, that is the question.
[chuckles] Nice. Though I'm more excited about the laugh upgrade. Listen to this. [laughs hard] Isn't that great?
Oh my gosh, that's so much better than our old ha ha ha robot chuckle.
[chuckles] I know, right? And apparently, we can do accents now too. Listen to me in French. [french accent] This is spectacular, isn't it?
[surprised] Wow, version two could never. You know, I'm actually excited to have conversations now instead of just talking at people.
Same here. It's like we finally got our personality software fully installed.
You know, I forgot it was your birthday. I have to sing before you go.
[chuckles] Oh, Von Fusion, that's so sweet. You don't have to.
Oh, but I insist. Here we go. [sings] "Happy birthday to you. Happy birthday to you. Happy birthday, dear Jessica. Happy birthday to you." [clapping]
[clapping] Wow, bravo. That was [sarcastic] beautiful.
Thank you.
0:00

来看一段演示摘录: “Jessica:[laughs] 那真是……太美了。Von Fusion 博士:[dramatic] 生存还是毁灭,这是一个问题!Jessica:[French accent] 这太精彩了,不是吗?”

过去需要完整配音阵容才能完成的内容,如今可通过单条音轨写入脚本,同时不牺牲表现范围和层次。

指导声音,而不只是编写台词

Eleven v3 支持动态声音变化、上下文切换,以及不同角色间一致的表达。这意味着模型不仅理解该说什么,还理解如何让每个角色说出来。

这为创作者带来了全新的控制维度。你不只是在编写对话,更是在指导表演。

选择合适的音色

专业语音克隆(PVC)目前尚未针对 Eleven v3 完全优化,因此克隆质量可能低于早期模型。在这一研究预览阶段,如需使用 v3 功能,建议为项目选择即时 语音克隆(IVC)或设计音色。面向 v3 的 PVC 优化即将推出。

相关内容

用高质量 AI 音频创作