推出 Eleven v4认识 Eleven v4:迄今情感表现最丰富的模型。 Creator+ 套餐含 3 倍点数优惠,截止至 10 月 12 日

跳至内容

Eleven v4 Audio Tags 列表:先试用,再自己编写

发布时间

收听收听本文

Eleven v4 是一款突破性的文本转语音模型,能将文本转化为表演。Audio Tags 可让你精细控制情绪表达,从而指导表演。加入 [whispers] 或 [excited] 这类方括号中的自然语言提示,听听 v4 如何相应调整演绎。

这份 ElevenLabs Audio Tags 列表涵盖完整的 Eleven v4 情绪范围,以及演绎、节奏、反应、口音和音效标签。掌握基础后,我们还会介绍如何用自然语言编写自己的标签。

探索 Eleven v4 的情绪

试听情绪范围
ExcitedPlayfulPeacefulAmazedAnxiousFrustratedScaredTiredExcitedPlayfulPeacefulAmazedAnxiousFrustratedScaredTiredExcitedPlayfulPeacefulAmazedAnxiousFrustratedScaredTiredExcitedPlayfulPeacefulAmazedAnxiousFrustratedScaredTired

摘要

  • Audio Tags 是方括号中的自然语言信号,Eleven v4 会将其读取为表演指示,从而改变后续文字的演绎方式。
  • Eleven v4 在 Artificial Analysis Speech Arena [September 2026] 中排名第一,Audio Tags 让你能够逐句指导表演。
  • Eleven v4 情绪范围转盘让你试听不同情绪,快速了解模型能力。
  • 标签并不限于固定列表;你可以组合特质来编写标签,例如 [whispering, fearful],也可以描述台词背后的情境或角色。
  • Audio Tags 可通过 ElevenAPI 在 Eleven v4、Eleven v4 Turbo 和 Eleven v3 上使用。

什么是 ElevenLabs Audio Tags?

Audio Tags 是可嵌入文本方括号内的自然语言提示。Eleven v4 会将其视为表演标记,作为改变特定单词或短语演绎方式的指示。你可以在 ElevenLabs 文本转语音应用中直接写入脚本,模型会将其生动呈现。

Eleven v4 是 Artificial Analysis Speech Arena 中排名第一的 TTS 模型,听众会在这里投票选出最自然的 文本转语音。1 Audio Tags 让你能精确指导每句台词的演绎方式,将表演提升到更高水平。

以下是在脚本中使用 Audio Tags 的方法:

  • 将标签放在它影响的文字之前:写下“[shouts] I can’t believe you said that to me”时,整句都会以喊叫方式演绎。
  • 情绪会持续:添加标签后,它的情绪会延续到整句,因此只有想改变演绎时才需要加入另一个标签。例如,“[proud] I’ve been cooking for over a decade. I know how to boil an egg. [startled] What’s that burning smell?”
  • 组合标签,叠加指示:在同一组方括号内用逗号分隔 Audio Tags,例如 [whispering, playful],可为 TTS 表演增添细腻层次。
  • 将标签与标点搭配:Audio Tags 设定情绪和演绎方式,而你可以在脚本中通过标点自然控制节奏。加入省略号、破折号或大写字母来塑造韵律。

想深入了解标签的工作方式及其如何取代 SSML,可阅读我们的Audio Tags 完整指南。

Text-to-speech leaderboard: Eleven v4 leads with 1319 Elo, ahead of Sonic 3.6. Take this further by selecting from the ElevenLabs Audio Tags list

情绪 Audio Tags 列表

Eleven v4 情绪转盘包含数十种情绪,每种都由 v4 演绎,让你在写下任何标签前先听出差别。

点击下方任意情绪即可在转盘中试听,或者直接尝试不同句子和情绪,感受 Eleven v4 如何呈现情绪 Audio Tags。

虽然这里展示的情绪只是少数,你可以用自然语言实现任何想要的情绪。下面的示例使用了上文未列出的多种 Audio Tags,展示自然语言提示的效果。

[jittery] Okay, final question of the pub quiz, and we're tied for first. [intrigued] "Which planet has the most moons?" Hmm. [smug] Easy. It's Saturn, everyone knows that. [puzzled] Wait, why is Priya shaking her head? [disgusted] Jupiter? You want to write down Jupiter? [frazzled] We only have ten seconds, just pick one, pick one! [hopeful] [nervous] Fine. Saturn. Hand it in. [long pause] [triumphant] YES! Saturn! We won! [mischievous] [slightly pause] Priya, I believe you owe me a drink.
0:00

此示例由 Jonathan Livingston 演绎。

演绎和音量 Audio Tags 列表

演绎和音量标签可控制台词的响度或强度,不受所附情绪影响。若希望 v4 更准确地遵循你的设想,可叠加演绎和情绪标签,实现更精细的控制。

以下是几个示例:

  • [whispers] 别动,它就在你身后。
  • [shouts] 所有人立刻撤离大楼!
  • [softly] 你已经尽力了。
  • [quietly] 我觉得他们已经走了。
  • [low, threatening] 你真不该来这里。

来看看音量和演绎控制的实际效果。

[hushed] Eighteenth hole. One putt to win the championship, and the crowd has gone completely silent. [barely audible] He's lining it up now. The ball is rolling... still rolling... [booming] IT'S IN! HE'S DONE IT! The championship is his, and this place has gone absolutely wild! [disbelief] Twenty years I've been calling golf, and I have never seen anything like that.
0:00

上面的示例由 Rod 演绎。

节奏 Audio Tags 列表

节奏 Audio Tags 可改变台词的语速。用它们营造悬念,或铺垫恰到好处的喜剧时刻。当时机和文字同样重要时,就该使用节奏标签。

额外提示:标点会自然改善文本转语音表现,因此可同时使用,以全面控制句子:

  • [slowly] 获胜者是……
  • [rushed] 抱歉,我迟到了,火车坏了,我一路跑过来的。
  • [pause] 接着电话响了。
  • [drawn out] 不会吧——
[slowly] Ten... nine... eight... seven... [rushed] Wait, wait, wait, hold the countdown, someone left their coffee on the console! [snappy] Can you get that out of here? [long pause] [drawn out] Okaaay, thank you. It's been moved. [excited] Resume the count! [speedy] Three... two... one... LIFTOFF!
0:00

Adam 为此示例配音。

拟人反应 Audio Tags 列表

反应标签可为句子加入笑声、喘气、哭泣、咳嗽和叹气等声音。它们能让文本更生动,让 TTS 音频示例听起来更自然,仿佛人即兴说话,而不是照着脚本念。

以下是一些反应 Audio Tags:

  • [laughs] 你居然真的上当了?
  • [sighs] 好吧,我去洗碗。
  • [gasps] 那是真钻石吗?
  • [clears throat] 请大家注意一下。
  • [crying] 我没想到你会回来。

在情绪需要时,Eleven v4 还会加入细微的人性化处理。试听情绪转盘,你会听到“Y- you came back?”和“Ugh, this is real?”之类的台词,其中的结巴或呻吟由 v4 自行加入。

[clears throat] Hi, everyone. For those who don't know me, I'm the best man. [laughs] Well, I'm the only man Tom could find at short notice. [sighs] When Tom first told me about Adam, I thought, there's no way he’s real. [gasps] Sorry, is that the cake? It's enormous. Anyway. [starts crying] I've never seen him this happy. [laughs] Okay, I'm fine. I'm fine. To Tom and Adam!
0:00

如需在自己的作品中使用此声音,请查找 Jack John。

口音和角色 Audio Tags 列表

借助 Eleven v4,使用不同口音或角色声音构建丰富场景比以往更轻松。只需少量 Audio Tags,就能在保留声音原有特质的同时,将其转变为新的角色人格。一次生成中,一个声音即可饰演整组角色。

  • [British accent] 来杯茶吗?
  • [French accent] 欢迎来到我的小咖啡馆。
  • [Australian accent] 别担心,伙计,我们会处理好的。
  • [pirate voice] 扬起船帆,递上朗姆酒。
Let's visit four stops in thirty seconds. First, London. [British accent] Mind the gap. Lovely weather we're having. [excited] Next, Dublin! [Irish accent] Grand day for it, isn't it? Sure, a bit of rain never hurt anyone. [rushed] No time, no time, on to Sydney! [Australian accent] G'day! Watch out for the seagulls, they'll nick your chips. [pirate voice] Arg, now the high seas, where the tour ends and the treasure begins!
0:00

Lauren 协助演绎了此示例。

音效 Audio Tags 列表

音效 Audio Tags 可直接在生成内容中加入非语音事件。例如,构建叙事场景或电子游戏音轨时,无需单独的 SFX 音轨也能增添戏剧效果。当然,如果需要特定效果,也可以使用 AI 音效生成器。

  • [thunder rumbling] 它越来越近了。
  • [footsteps] 有人正在上楼。
  • [door creaking] 你好?有人在家吗?
  • [clapping] 谢谢,谢谢,你们太客气了。
[owl hooting] [nervous] Did you hear that? [gulp] It's just an owl, right? [twig snapping] [nervous] Okay, owls don't do that. [many footsteps] [scared] Something's walking around the tent. [zipper opening] [startled] Wait, who just opened the tent? [dog barking] [laughs] Biscuit! You scared me half to death. [sighs] Fine, you can sleep in here too.
0:00

上方片段中的声音是 Siren。

用自然语言编写自己的 Audio Tags

上面展示的任何 Audio Tags 都是很好的起点。但 ElevenLabs TTS 的妙处在于,你可以用自然语言编写任何新的 Audio Tag,为模型提供额外可用的上下文。

或者,如果没有单个词标签能表达你的需求,可以写出更完整的指示。

  • 组合情绪和特质: [tense, cautious] 或 [whispering, fearful]
  • 描述表达方式: [like a sports commentator, speeding up]
  • 描述情境: [out of breath after running up the stairs]
  • 描述角色: [a tired detective who has heard it all before]
  • 描述转变: [starting calm, then losing patience]
[hushed and reverent, like a nature documentary narrator] Here, in the quiet of the office kitchen, a rare creature emerges. [barely containing excitement] The intern. [slow and suspenseful] He approaches the last slice of birthday cake. He has been waiting for this moment all afternoon. [speeding up, like a sports commentator] He's going for it, he's reaching, he's almost there, he's- [sudden, crushed disappointment] Someone from accounts got there first. [hushed and reverent] Nature, as always, is cruel.
0:00

使用 Spuds Oxley,让场景生动起来。

选择 Audio Tags 的技巧

用自然语言制作文本转语音音频示例非常灵活,但要获得理想效果,可能也需要反复调整。

为让 Eleven v4 的 TTS 使用体验更顺畅,以下是使用 Audio Tags 的一些技巧:

  • 先试听,再编写:在 情绪转盘中播放几种情绪,了解 v4 如何诠释每一种情绪(也可借此获取灵感),再为台词选择最接近的效果。
  • 先换标签,再重写台词:如果演绎不理想,尝试相邻情绪,例如用 [let down] 替代 [despair],然后重新生成。
  • 每个分句使用一个标签:同一段文字搭配对比鲜明的标签可能会让表演模糊。在希望情绪变化的位置加入新标签。
  • 为模型提供完整场景:Eleven v4 在有上下文时表现更好。文本转语音应用每次生成最多支持 10,000 个字符,你有足够空间构建层层递进的指示,为场景创造出色效果。
  • 先用预设,再具体说明。如果 [nervous] 接近但不完全合适,请写出缺少的感觉:[nervous, trying to sound confident]。

但最重要的是,直接进入 ElevenLabs TTS 应用进行实验,才是快速上手并开始制作高质量 TTS 音频的最佳方式。

Guide to six ElevenLabs audio tag list types and layering multiple cues with comma-separated tags.

使用 Eleven v4 开始创作沉浸式音频

不存在唯一权威的 ElevenLabs Audio Tags 列表,因为你可以用自然语言创建任何想要的组合。此列表中的每个标签都适用于 Eleven v4,而你在编写场景时想到的标签也同样适用。

从 声音库的 17,500+ 个声音中选择一个,粘贴脚本,加入第一个 Audio Tag,即可开始。

了解更多关于 Eleven v4,或注册,创作出色的音频表演。

ElevenLabs Audio Tags 常见问题

  1. Artificial Analysis,Provider Voice Arena Preference Elo,2026 年 9 月 30 日

相关内容

用高质量 AI 音频创作