Last refreshed 2026-06-29. Next refresh: weekly.
Why use Scribe v2 Realtime on ElevenLabs API?
ElevenLabs API offers Scribe v2 Realtime with competitive pricing. ElevenLabs' API provides hosted text-to-speech, speech-to-speech, dubbing, and voice models for creative production and realtime conversational applications.
Input / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: scribe_v2_realtimeModel ID
scribe_v2_realtimeRequest example
Curated snippets for this provider are not sourced yet. Use ElevenLabs API documentation with model ID
scribe_v2_realtime.Gotchas
- Use provider model ID "scribe_v2_realtime", not the LLMReference slug "scribe-v2-realtime".
Capabilities
Audio
About Scribe v2 Realtime
Scribe v2 Realtime is ElevenLabs' streaming speech-to-text model for voice agents, released November 2025. Delivers ~150ms latency with WebSocket streaming across 90+ languages. Supports Voice Activity Detection (VAD), manual commit control, and the same language coverage as Scribe v2. Priced at $0.39/hr of audio. API ID: scribe_v2_realtime.
Get Started
Model Specs
Released2025-11-01
ArchitectureAudio / Speech