Eleven v4

Expressive multilingual voiceover with audio tags and separate stability and voice similarity controls. Supports up to 2000 characters per generation.

Cost

from 8 tokens

ImageProvider: ElevenLabs

Try Eleven v4 in Givon AI

Open a working project for regular creation, or use Playground to inspect the model parameters.

Use through API

Request fields, constraints, and ready-to-use examples for developers. The user workflow starts above.

Run generation
IDEMPOTENCY_KEY="${IDEMPOTENCY_KEY:-$(uuidgen)}"
curl -X POST https://api.givon.ai/api/v1/generations \
  -H "Authorization: Bearer $GIVON_API_KEY" \
  -H "Idempotency-Key: $IDEMPOTENCY_KEY" \
  -H "Content-Type: application/json" \
  -d '{"type":"image","model":"elevenlabs-tts-v4","input":{"prompt":"studio portrait of a red corgi, soft light","stability":0.5,"similarity":0.75}}'

Input fields

* required
prompt*prompt

Voiceover script/text to speak.

Type
string
Default
—
Allowed
up to 2000 chars
voice*voice

Voice id from GET /api/v1/voices.

Type
string
Default
—
Allowed
string
stabilitystability

Controls consistency of delivery.

Type
number
Default
0.5
Allowed
from 0 · up to 1 · step 0.01
similaritysimilarity

Controls how closely the output follows the selected voice.

Type
number
Default
0.75
Allowed
from 0 · up to 1 · step 0.01

Cost

from 8 tokens
Chars- 0- 250default8 tk
Chars- 251- 50014 tk
Chars- 501- 100024 tk
Chars- 1001-plus44 tk

The variant is selected automatically from request fields, so you do not need to send it.

Capabilities

Modestext_to_speech
Get API key