Model chooser / Transcribe audio, calls and meetings / One-off
Scribe v2 has one of the lowest word error rates on Artificial Analysis's speech-to-text benchmark from a major vendor, labels who is speaking, and runs in a web app where you upload a file and download the transcript.
Why ElevenLabs Scribe v2
How to use it
Prompt for summarising the transcript
Edit the parts in capitals, then run it.
Below is a transcript of a CALL TYPE between SPEAKER A (us) and SPEAKER B (them). Give me: 1. Their situation and the problem they described, in their words where possible 2. Objections or concerns they raised, each with a direct quote 3. What was agreed and who owns each next step 4. Anything they said that contradicts what is in our CRM: PASTE CRM NOTES Quote exactly. If something is ambiguous in the transcript, say so. TRANSCRIPT: PASTE HERE
Alternatives that also work
Pick it when the recording is confidential and must stay on your machine. It is the best open model on the Hugging Face Open ASR leaderboard (4.31 average WER).
Pick it when you want the transcript and the summary in one step. It takes audio directly.
Pick it when you need the transcript in seconds rather than minutes. It is the fastest model on Artificial Analysis's benchmark.
Watch out for
Sources
Checked . Models change monthly; we re-check this page when they do.
Wiring it into a workflow that runs every week, with evals, fallbacks and a cost you can predict, is the work. Fifteen minutes, no deck.