LLM Reference
fal API

Using Eleven Flash v2.5 on fal API

Implementation guide · ElevenLabs Text-to-Speech · ElevenLabs

Serverless

Quick Start

  1. 1
    Create an account at fal API and generate an API key.
  2. 2
    Use the fal API SDK or REST API to call fal-ai/elevenlabs/tts/turbo-v2.5 — see the documentation for request format.

Code Examples

See fal API documentation for integration details.

About fal API

fal serves asynchronous and real-time model APIs on its infrastructure with queue handling, automatic scaling, SDK and HTTP access, and model-specific output-based billing. Its public catalog includes first-party and partner-hosted endpoints across image, video, audio, vision, and 3D generation; model cards expose the applicable schema, pricing, and license information.

fal API is a developer platform for calling a large catalog of production-ready image, video, audio, vision, 3D, and multimodal generation models through a unified API surface.

Pricing on fal API

Capabilities

Audio

About Eleven Flash v2.5

Eleven Flash v2.5 is ElevenLabs' ultra-low-latency TTS model for real-time applications and voice agents, released December 18, 2024. Achieves ~75ms latency with 32 languages (adds Hungarian, Norwegian, Vietnamese over Flash v2). Maximum input: 40,000 characters per request. Priced at $0.05/1K characters ($50/1M chars) — 50% cheaper than Multilingual v2. API ID: eleven_flash_v2_5.

Model Specs

Released2024-12-18
ArchitectureAudio / Speech

Provider

fal API
fal API

fal

Paris, Île-de-France, France