Skip to main content

Overview

ElevenLabs provides the text-to-speech models in varg. Multiple model variants are available balancing quality, speed, and language support.

Quick start

Available voices

Any valid ElevenLabs voice_id works. The names above are convenience aliases for built-in voices. Browse more at ElevenLabs Voice Library.

Parameters

string
required
The text to convert to speech.
string
default:"rachel"
Voice name or ElevenLabs voice ID.
string
default:"eleven_multilingual_v2"
Speech model variant (see table above).
object
stability — 0 to 1 (default 0.5). Higher = more consistent, lower = more expressive. similarity_boost — 0 to 1 (default 0.75). How closely to match the voice.

Choosing a model

Composition example

Use speech in a video composition with captions:

Pricing

VEED Fabric

Animate a portrait with generated speech.

Sync Lipsync

Apply speech to existing video.

Whisper

Transcribe audio back to text.