Scribe V2
/v1/runScribe V2 is ElevenLabs's speech recognition model. Transcribe audio content with industry-leading accuracy across multiple languages and accents.
Request body
Submit an async generation request. The model field selects the model; other fields are model-specific input parameters.
Model identifier. Set to elevenlabs/scribe-v2.
Default: elevenlabs/scribe-v2
URL of the audio file to transcribe
Tag audio events like laughter, applause, etc.
Default: true
Whether to annotate who is speaking
Default: true
Language code of the audio
Response Schema
The submit endpoint returns an accepted generation task. Poll the result endpoint with the returned id for terminal outputs or errors.
Error message if the task failed. Empty on success.
Unique identifier for the generation task.
Model ID used for the prediction.
Array of generated content. Empty when status is not completed.
Status of the task: pending, running, completed, failed, or timeout.
Allowed values: pending, running, completed, failed, timeout
Model capabilities
Capabilities declared by the model registry.
Default: speech-to-text
Execution mode declared by the model registry.
Default: async

