Last refreshed 2026-06-29. Next refresh: weekly.
Why use Scribe v2 on ElevenLabs API?
ElevenLabs API offers Scribe v2 with competitive pricing. ElevenLabs' API provides hosted text-to-speech, speech-to-speech, dubbing, and voice models for creative production and realtime conversational applications.
Input / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: scribe_v2Model ID
scribe_v2Request example
Curated snippets for this provider are not sourced yet. Use ElevenLabs API documentation with model ID
scribe_v2.Gotchas
- Use provider model ID "scribe_v2", not the LLMReference slug "scribe-v2".
Capabilities
Audio
About Scribe v2
Scribe v2 is ElevenLabs' current state-of-the-art batch speech-to-text model, released January 12, 2026. Improvements over v1 for long-form audio, extended silences, and tone changes. Supports 90+ languages, word-level timestamps, 32-speaker diarization, 56 entity types, and keyterm prompting (up to 1,000 terms). Base pricing: $0.22/hr; entity detection add-on: $0.07/hr; keyterm prompting add-on: $0.05/hr. API ID: scribe_v2.
Get Started
Model Specs
Released2026-01-12
ArchitectureAudio / Speech