Eleven v4 and Eleven v4 Turbo

Eleven v4 and Eleven v4 Turbo are now available. Eleven v4 provides higher-quality, more expressive speech and improved voice cloning across more than 90 languages. Eleven v4 Turbo provides the same model family for real-time applications, with median inference latency of approximately 100 ms.

Use eleven_v4 through the Text to Dialogue API for content creation and long-form audio. Use eleven_v4_turbo through the Text to Dialogue WebSocket for agents and interactive applications.

Speech to Text

  • Transcript editing: Batch and realtime transcription now accept a natural-language edit instruction of up to 2,000 characters. Batch responses return the result in edited_transcript, while realtime sessions emit an edited_transcript event. See the batch and realtime guides for incompatibilities, response handling and pricing.

ElevenAgents

  • Agent models: Added gpt-6-sol, gpt-6-luna, glm-52, deepseek-v41-flash, claude-opus-5 and claude-opus-5-5 to the supported agent LLM options.
  • Agent test usage: Test invocation summaries can include total credits and USD price. Individual test runs can include credits and a charging breakdown.

Flows

  • Template API: Added endpoints to list and retrieve published Flows templates, start template runs and retrieve run results. Runs can target a published version and send their terminal result to a configured webhook.

Text to Dialogue

  • Generation continuity: Text to Dialogue requests can use preceding or following text and request IDs to maintain continuity across generations.
  • Professional voice mode: Text to Dialogue requests can use the Instant Voice Clone version of a Professional Voice Clone to reduce latency or improve expressiveness.

SDK releases

JavaScript SDK

  • v2.70.0 - Added batch and realtime Speech to Text transcript editing types, branchId filtering and cost fields for test invocation summaries, knowledge base auto-discovery, Text to Dialogue usePvcAsIvc and the deepseek-v41-flash agent model.

Python SDK

  • v2.70.0 - Added batch and realtime Speech to Text transcript editing types, branch_id filtering and cost fields for test invocation summaries, knowledge base auto-discovery, Text to Dialogue use_pvc_as_ivc and the deepseek-v41-flash agent model.

ElevenLabs CLI

  • v1.4.0 - Fixed requests with an explicit API key so they no longer also send stored OAuth credentials, which caused authentication failures. Added descriptions for the feedback command group and instructions for reporting missing capabilities to generated agent skills. Regenerated the API client with Flows template endpoints and current schema types.

API

New endpoints

Flows

  • List templates - GET /v1/flows/templates - List published templates and their runnable versions.
  • Get template - GET /v1/flows/templates/{template_id} - Retrieve a template and its runnable versions.
  • Create template run - POST /v1/flows/templates/{template_id}/runs - Start a template run with input values, an optional version and an optional webhook.
  • List template runs - GET /v1/flows/templates/{template_id}/runs - List template runs with cursor pagination and optional version filtering.
  • Get template run - GET /v1/flows/templates/{template_id}/runs/{run_id} - Retrieve a template run and its output status.

Updated endpoints

Speech to Text

  • Convert speech to text - POST /v1/speech-to-text
    • Added optional nullable transcript_edit (string), which accepts a natural-language edit instruction of up to 2,000 characters.
    • Responses add optional nullable edited_transcript. A successful edit contains required kind: "transcript" and text fields. A failed edit contains required kind: "error", error_type: "edit_failed" and message fields.
  • Realtime Speech to Text accepts transcript_edit (string) and emits edited_transcript events with required text and edited_text fields.

ElevenAgents

  • Create agent, Get agent and Update agent
    • Added gpt-6-sol, gpt-6-luna, glm-52, deepseek-v41-flash, claude-opus-5 and claude-opus-5-5 to the LLM enum.
    • Alerting monitors add optional nullable enabled (boolean) to control whether a monitor can send notifications.
  • List conversations and Text search
    • data_collection_params and dynamic_variable_params add neq and in operators. Values for in use a pipe delimiter.
    • evaluation_params accepts success, failure or unknown as the result value.
  • Smart search - GET /v1/convai/conversations/messages/smart-search
    • Added optional nullable branch_id (string).
  • List test invocations - GET /v1/convai/test-invocations
    • Added optional nullable branch_id (string).
    • Test invocation summaries add optional nullable credits_used (integer) and total_price (number).
  • Get test invocation - GET /v1/convai/test-invocations/{test_invocation_id}
    • Individual test runs add optional nullable credits_used (integer) and charging (ConversationChargingCommonModel).
  • Simulate conversation and Simulate conversation stream
    • Changed the optional new_turns_limit default from 10,000 to 100 and set its maximum to 100.
  • Create crawl job and Get crawl job
    • Added optional auto_discover (boolean, default false) to follow links and add newly discovered pages during auto-sync. This requires enable_auto_sync: true.
  • Refresh knowledge base document - POST /v1/convai/knowledge-base/{documentation_id}/refresh
    • AutoSyncInfo adds optional auto_discover (boolean, default false) to indicate whether refreshes crawl and add newly discovered pages.
  • List MCP server tools - GET /v1/convai/mcp-servers/{mcp_server_id}/tools
    • Responses add optional tool_approval_statuses (array). Each item contains required tool_id (string) and state (up_to_date, needs_review or not_approved), with optional approval policy and approved-definition fields.

Text to Dialogue

  • Convert text to dialogue, Stream text to dialogue, Convert with timestamps and Stream with timestamps
    • Added optional use_pvc_as_ivc (boolean, default false) to use the Instant Voice Clone version of a Professional Voice Clone.
    • Added optional nullable previous_text and future_text (strings, maximum 100 characters) and previous_request_ids and next_request_ids (string arrays, maximum 3 IDs) for generation continuity.
    • The optional settings object adds nullable similarity (number, 0–1, default 0.75).

Text to Speech

Flows

Voice Design

Workspaces

  • Create service account API key - POST /v1/service-accounts/{service_account_user_id}/api-keys
    • Added optional nullable tts_concurrency_limit, music_concurrency_limit and dubbing_concurrency_limit (integers) for enterprise service account API keys.
  • Update service account API key - PATCH /v1/service-accounts/{service_account_user_id}/api-keys/{api_key_id}
    • Added optional tts_concurrency_limit, music_concurrency_limit and dubbing_concurrency_limit. Each field accepts an integer, clear, no_update or null and defaults to no_update.

Models

  • List models - GET /v1/models
    • Deprecated required max_characters_request_free_user and max_characters_request_subscribed_user. These fields are not enforced; use maximum_text_length_per_request instead.