Eleven v4 API

elevenlabs/eleven_v4
ElevenLabs v4 is the most expressive ElevenLabs text-to-speech model, with inline audio tags for emotion and delivery in 90+ languages.
Output
Released
Sep 28, 2026

How to use Eleven v4 API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to elevenlabs/eleven_v4.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/tts",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "elevenlabs/eleven_v4",
      "text": "Hello from AI/ML API",
      "voice": "Rachel"
    },
)
print(r.json()["audio"])  # URL to the generated audio
const r = await fetch("https://api.aimlapi.com/v1/tts", {
  method: "POST",
  headers: {
    Authorization: "Bearer " + process.env.AIMLAPI_KEY,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "elevenlabs/eleven_v4",
    "text": "Hello from AI/ML API",
    "voice": "Rachel"
  }),
});
console.log((await r.json()).audio);
curl -X POST https://api.aimlapi.com/v1/tts \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"elevenlabs/eleven_v4","text":"Hello from AI/ML API","voice":"Rachel"}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Eleven v4 API Pricing

TypePrice
Input
$28.6 / 1M characters
Output

Billed per character of input text.

Eleven v4 Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
TTS Arena
1321
Human preference Elo from blind pairwise listening comparisons of speech samples, measured independently by Artificial AnalysisSourceOctober 6, 2026

Eleven v4 vs other models

ModelInputOutputContextBest for
Eleven v4
This page
$28.6 / 1M characters
Speech synthesis
$78 / 1M characters
Speech synthesis
$130 / 1M characters
Speech synthesis
$28.6 / 1M
$28.6 / 1M
Speech synthesis

Frequently asked questions

Eleven v4 is ElevenLabs' text-to-speech model, described by the vendor as its most emotive and highest quality speech synthesis model. It turns text into spoken audio.

$28.6 per 1M characters through AI/ML API. Speech models are billed by characters of input text rather than by tokens.

Text in, audio out. You pass the text and a voice, and the response carries a URL to the generated audio.

Both are ElevenLabs' v4 generation, with the same expressive voices and inline audio tags in 90+ languages. v4 is positioned for highest quality, v4 Turbo for real-time use such as agents and live conversation. Turbo is half the price at $14.3 per 1M characters.

The text-to-speech endpoint, POST /v1/tts, with model set to elevenlabs/eleven_v4.

Start building with Eleven v4

Get API Key
1000+ models, one API.