LLM Reference

Deepgram Nova-2

Released
2023-09-19
Last refreshed
2026-06-29
Status
Researched 89d ago
ProprietaryCommercial use: conditionalMultimodalVisionAudio

Deepgram Nova-2 is worth evaluating for vision when its provider route and context window match the workload.

Use it for

  • Teams evaluating vision
  • Buyers comparing 1 tracked provider route

Do not use it for

  • Strict JSON or tool-calling flows
Specifications
Released
2023-09-19
Architecture
Audio / Speech
Specialization
speech-recognition
Openness
Proprietary
License
ProprietaryCommercial use: conditional
Weights
Not released
Code
Unknown
Training
Pretrained
Created by

AI-powered speech recognition and text-to-speech platform.

San Francisco, USA
Founded 2014
Website
Pricing
Output / 1M
-
Input / 1M
-

Cheapest of 1 route · Deepgram API

About

Nova-2 is Deepgram's previous-generation flagship speech-to-text model, released September 2023. Delivers ~36% WER improvement over Whisper Large across tested domains (8.4% median WER), with improved entity recognition, punctuation, and capitalization. Supports 36+ languages and 10 domain-specific variants (general, meeting, phonecall, voicemail, finance, conversationalai, video, medical, drivethru, automotive). Batch: $0.0043/min; streaming: $0.0077/min. API ID: nova-2.

Top use-case fit

Vision

Included by capability and metadata signals in the decision map.

Provider price ladder

Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
Deepgram API--
ServerlessPartial

Capabilities

MultimodalAudio

Benchmark peer barsfor Vision

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.

API versions

nova-2nova-2-general