LLM Reference

Scribe v2

Released
2026-01-12
Last refreshed
2026-06-29
Status
Researched 89d ago
ProprietaryCommercial use: conditional

Scribe v2 is worth evaluating for general LLM work when its provider route and context window match the workload.

Use it for

  • Teams evaluating general LLM work
  • Buyers comparing 1 tracked provider route

Do not use it for

  • Vision or document-understanding workloads
  • Strict JSON or tool-calling flows
Specifications
Released
2026-01-12
Architecture
Audio / Speech
Specialization
speech-recognition
Openness
Proprietary
License
ProprietaryCommercial use: conditional
Weights
Not released
Code
Unknown
Training
Pretrained
Created by

AI audio research and text-to-speech platform.

New York, New York, United States
Founded 2022
Website
Pricing
Output / 1M
-
Input / 1M
-

Cheapest of 1 route · ElevenLabs API

About

Scribe v2 is ElevenLabs' current state-of-the-art batch speech-to-text model, released January 12, 2026. Improvements over v1 for long-form audio, extended silences, and tone changes. Supports 90+ languages, word-level timestamps, 32-speaker diarization, 56 entity types, and keyterm prompting (up to 1,000 terms). Base pricing: $0.22/hr; entity detection add-on: $0.07/hr; keyterm prompting add-on: $0.05/hr. API ID: scribe_v2.

Top use-case fit

No primary decision-task fit is mapped for this model yet.

Provider price ladder

Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
ElevenLabs API--
ServerlessPartial

Capabilities

Audio

Benchmark peer barsfor Coding

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.

API versions

scribe_v2