> This is a page from the ElevenLabs documentation. For a complete page index, fetch https://elevenlabs.io/docs/llms.txt. For the full documentation in a single file, fetch https://elevenlabs.io/docs/llms-full.txt. # Get transcript GET https://api.elevenlabs.io/v1/speech-to-text/transcripts/{transcription_id} Retrieve a previously generated transcript by its ID. Reference: https://elevenlabs.io/docs/api-reference/speech-to-text/get ## Servers - `https://api.elevenlabs.io` (Production, default) - `https://api.us.elevenlabs.io` (Production US) - `https://api.eu.residency.elevenlabs.io` (Production EU) - `https://api.in.residency.elevenlabs.io` (Production India) - `https://api.sg.residency.elevenlabs.io` (Production Singapore) ## Request ### Path parameters - `transcription_id` (string, required) — The unique ID of the transcript to retrieve ## Response ### 200 The transcript data - `speech_to_text_transcripts_get_Response_200` ## Errors ### 422 Unprocessable Entity Error Validation Error - `detail` (list of ValidationError, optional) ## Types ### SpeechToTextChunkResponseModel Chunk-level detail of the transcription with timing information. - `language_code` (string, required) — The detected language code (e.g. 'eng' for English). - `language_probability` (double, required) — The confidence score of the language detection (0 to 1). - `text` (string, required) — The raw text of the transcription. - `words` (list of SpeechToTextWordResponseModel, required) — List of words with their timing information. - `channel_index` (integer, optional, nullable) — The channel index this transcript belongs to (for multichannel audio). - `additional_formats` (list of AdditionalFormatResponseModel, optional, nullable) — Requested additional formats of the transcript. - `transcription_id` (string, optional, nullable) — The transcription ID of the response. - `entities` (list of DetectedEntity, optional, nullable) — List of detected entities with their text, type, and character positions in the transcript. - `audio_duration_secs` (double, optional, nullable) — The duration of the audio that was transcribed in seconds. - `edited_transcript` (SpeechToTextChunkResponseModelEditedTranscript, optional, nullable) — Result of the optional transcript edit: the edited text, or an error if it could not be produced. Absent when no edit was requested. ### MultichannelSpeechToTextResponseModel Response model for multichannel speech-to-text transcription. - `transcripts` (list of SpeechToTextChunkResponseModel, required) — List of transcripts, one for each audio channel. Each transcript contains the text and word-level details for its respective channel. - `transcription_id` (string, optional, nullable) — The transcription ID of the response. - `audio_duration_secs` (double, optional, nullable) — The duration of the audio that was transcribed across all channels in seconds. ### ValidationError - `loc` (list of ValidationErrorLocItems, required) - `msg` (string, required) - `type` (string, required) ### SpeechToTextWordResponseModel Word-level detail of the transcription with timing information. - `text` (string, required) — The word or sound that was transcribed. - `type` (enum, required) — The type of the word or sound. 'audio_event' is used for non-word sounds like laughter or footsteps. - Allowed values: `word`, `spacing`, `audio_event` - `logprob` (double, required) — The log of the probability with which this word was predicted. Logprobs are in range [-infinity, 0], higher logprobs indicate a higher confidence the model has in its predictions. - `start` (double, optional, nullable) — The start time of the word or sound in seconds. - `end` (double, optional, nullable) — The end time of the word or sound in seconds. - `speaker_id` (string, optional, nullable) — Unique identifier for the speaker of this word. - `characters` (list of SpeechToTextCharacterResponseModel, optional, nullable) — The characters that make up the word and their timing information. - `channel_index` (integer, optional, nullable) — The channel this word was spoken on (for multichannel audio). Null for single-channel transcriptions. ### AdditionalFormatResponseModel - `requested_format` (string, required) — The requested format. - `file_extension` (string, required) — The file extension of the additional format. - `content_type` (string, required) — The content type of the additional format. - `is_base64_encoded` (boolean, required) — Whether the content is base64 encoded. - `content` (string, required) — The content of the additional format. ### DetectedEntity An entity detected within transcribed text. - `text` (string, required) — The text that was identified as an entity. - `entity_type` (string, required) — The type of entity detected (e.g., 'credit_card', 'email_address', 'person_name'). - `start_char` (integer, required) — Start character position in the transcript text. - `end_char` (integer, required) — End character position in the transcript text. ### SpeechToTextChunkResponseModelEditedTranscript Result of the optional transcript edit: the edited text, or an error if it could not be produced. Absent when no edit was requested. - `kind`: `error` (TranscriptEditError) - `error_type` ("edit_failed", required) — edit_failed: the edit could not be produced. - `message` (string, required) — A short, user-facing explanation of the failure. - `kind`: `transcript` (EditedTranscript) - `edited_text` (string, required) — The edited transcript text. If no edits were made it will be identical to the `text` field. - `text` (string, required) — The committed transcript text the edit instruction was applied to. - `message_type` (string, optional, default: edited_transcript) ### ValidationErrorLocItems ### SpeechToTextCharacterResponseModel - `text` (string, required) — The character that was transcribed. - `start` (double, optional, nullable) — The start time of the character in seconds. - `end` (double, optional, nullable) — The end time of the character in seconds. ## Examples **Response** ```json { "language_code": "en", "language_probability": 0.98, "text": "Hello world!", "words": [ { "end": 0.5, "logprob": -0.124, "speaker_id": "speaker_1", "start": 0, "text": "Hello", "type": "word" }, { "end": 0.5, "logprob": 0, "speaker_id": "speaker_1", "start": 0.5, "text": " ", "type": "spacing" }, { "end": 1.2, "logprob": -0.089, "speaker_id": "speaker_1", "start": 0.5, "text": "world!", "type": "word" } ] } ``` **SDK Code** ```python import requests url = "https://api.elevenlabs.io/v1/speech-to-text/transcripts/transcription_id" response = requests.get(url) print(response.json()) ``` ```javascript const url = 'https://api.elevenlabs.io/v1/speech-to-text/transcripts/transcription_id'; const options = {method: 'GET'}; try { const response = await fetch(url, options); const data = await response.json(); console.log(data); } catch (error) { console.error(error); } ``` ```go package main import ( "fmt" "net/http" "io" ) func main() { url := "https://api.elevenlabs.io/v1/speech-to-text/transcripts/transcription_id" req, _ := http.NewRequest("GET", url, nil) res, _ := http.DefaultClient.Do(req) defer res.Body.Close() body, _ := io.ReadAll(res.Body) fmt.Println(res) fmt.Println(string(body)) } ``` ```ruby require 'uri' require 'net/http' url = URI("https://api.elevenlabs.io/v1/speech-to-text/transcripts/transcription_id") http = Net::HTTP.new(url.host, url.port) http.use_ssl = true request = Net::HTTP::Get.new(url) response = http.request(request) puts response.read_body ``` ```java import com.mashape.unirest.http.HttpResponse; import com.mashape.unirest.http.Unirest; HttpResponse response = Unirest.get("https://api.elevenlabs.io/v1/speech-to-text/transcripts/transcription_id") .asString(); ``` ```php request('GET', 'https://api.elevenlabs.io/v1/speech-to-text/transcripts/transcription_id'); echo $response->getBody(); ``` ```csharp using RestSharp; var client = new RestClient("https://api.elevenlabs.io/v1/speech-to-text/transcripts/transcription_id"); var request = new RestRequest(Method.GET); IRestResponse response = client.Execute(request); ``` ```swift import Foundation let request = NSMutableURLRequest(url: NSURL(string: "https://api.elevenlabs.io/v1/speech-to-text/transcripts/transcription_id")! as URL, cachePolicy: .useProtocolCachePolicy, timeoutInterval: 10.0) request.httpMethod = "GET" let session = URLSession.shared let dataTask = session.dataTask(with: request as URLRequest, completionHandler: { (data, response, error) -> Void in if (error != nil) { print(error as Any) } else { let httpResponse = response as? HTTPURLResponse print(httpResponse) } }) dataTask.resume() ``` > ElevenLabs provides APIs and SDKs for text to speech, voice cloning, speech to text, sound effects, voice isolator, voice changer, and conversational AI agents. Build voice-enabled applications with lifelike audio generation.