Create a transcript from a media file that is accessible via a URL.
Authentication
API Key authentication via header
Request
Params to create a transcript
audio_url string Required
format: "url"
The URL of the audio or video file to transcribe.speech_models list of strings Required
List multiple speech models in priority order, allowing our system to automatically route your audio to the best available option. See Model Selection for available models and routing behavior.audio_end_at integer Optional
The point in time, in milliseconds, to stop transcribing in your media file. See Set the start and end of the transcript for more details.audio_start_from integer Optional
The point in time, in milliseconds, to begin transcribing in your media file. See Set the start and end of the transcript for more details.auto_highlights boolean Optional Defaults to
false
Enable Key Phrases, either true or falsecontent_safety boolean Optional Defaults to
false
Enable Content Moderation, can be true or falsecontent_safety_confidence integer Optional
25-100Defaults to50
The confidence threshold for the Content Moderation model. Values must be between 25 and 100.custom_spelling list of objects Optional
Customize how words are spelled and formatted using to and from values. See Custom Spelling for more details.disfluencies boolean Optional Defaults to
false
Transcribe Filler Words, like “umm”, in your media file; can be true or falsedomain string or null Optional
Enable domain-specific transcription models to improve accuracy for specialized terminology. Set to"medical-v1"to enable Medical Mode for improved accuracy of medical terms such as medications, procedures, conditions, and dosages.entity_detection boolean Optional Defaults to
false
Enable Entity Detection, can be true or falsefilter_profanity boolean Optional Defaults to
false
Filter profanity from the transcribed text, can be true or false. See Profanity Filtering for more details.format_text boolean Optional Defaults to
true
Enable Text Formatting, can be true or falseiab_categories boolean Optional Defaults to
false
Enable Topic Detection, can be true or falsekeyterms_prompt list of strings Optional
Improve accuracy with up to 200 (for Universal-2) or 1000 (for Universal-3 Pro) domain-specific words or phrases (maximum 6 words per phrase). See Keyterms Prompting for more details.language_code enum or null Optional
The language of your audio file. Possible values are found in Supported Languages. The default value is ‘en_us’.language_detection boolean Optional Defaults to
false
Enable Automatic language detection, either true or false.multichannel boolean Optional Defaults to
false
Enable Multichannel transcription, can be true or false.prompt string Optional
Provide natural language prompting of up to 1,500 words of contextual information to the model. See the Prompting Guide for best practices.
Note: This parameter is only supported for the Universal-3 Pro model.punctuate boolean Optional Defaults to
true
Enable Automatic Punctuation, can be true or false.redact_pii boolean Optional Defaults to
false
Redact PII from the transcribed text using the Redact PII model, can be true or false. See PII Redaction for more details.webhook_url string Optional
format: "url"
The URL to which we send webhook requests.
Response
Transcript created and queued for processing
audio_url string
format: "url"
The URL of the media that was transcribed.auto_highlights boolean
Whether Key Phrases is enabled, either true or false.id string
format: "uuid"
The unique identifier of your transcript.language_confidence double
0-1
The confidence score for the detected language, between 0.0 (low confidence) and 1.0 (high confidence). See Automatic Language Detection for more details.status enum
The status of your transcript. Possible values are queued, processing, completed, or error.
Allowed values: queued, processing, completed, error.webhook_auth boolean
Whether webhook authentication details were provided.