AI TraceTrace Foundation, Inc.
Productivity AutomationVerified

Reviewed and published by trentmaziarz, April 22, 2026. Discovered and drafted by our automated research pipeline.

ElevenLabs offers Scribe, an AI speech-to-text model that converts audio and video recordings into accurate, structured text transcripts in 99 languages, with features including speaker labeling, word-level timestamps, and audio-event tagging. Scribe launched in December 2024, with Scribe v2 following in January 2026.

Details

Scribe takes audio or video files as input and produces structured JSON transcripts as output, including speaker diarization (identifying who said what), character-level timestamps, and tagging of non-speech audio events such as laughter or applause. A real-time version, Scribe v2 Realtime, processes live speech with approximately 150 milliseconds of latency and is designed for use in conversational AI agents and meeting assistants. The tool is available via the web dashboard and API.

Products affected

ScribeScribe v2Scribe v2 RealtimeElevenLabs StudioElevenLabs API

Sources & Evidence

Cite this record

Trace Foundation. (2026). ElevenLabs: ElevenLabs offers Scribe, an AI speech-to-text model that converts audio and video recordings into accurate, structured text transcripts in 99 languages, with features including speaker labeling, word-level timestamps, and audio-event tagging. Scribe launched in December 2024, with Scribe v2 following in January 2026 (data as of 2026-04-22) [Data set record]. AI Trace. https://www.aitrace.org/r/practice/8c0319e4-a5eb-458e-8f31-bc6e1b6e5f05. Accessed September 11, 2026.

Stable link
https://www.aitrace.org/r/practice/8c0319e4-a5eb-458e-8f31-bc6e1b6e5f05
Data as of
April 22, 2026
Last verified
August 4, 2026

How to cite AI Trace

Other practices by ElevenLabs

Have evidence about ElevenLabs's AI practices? Submit a report.

Report a Sighting →