Speech 2.8 HD API

MiniMax Speech 2.8 — text-to-speech model for generating natural-sounding speech from text. Supports multiple quality modes (turbo and hd) optimized for latency or audio fidelity.
Output
$130 / 1M tokens
Released
Apr 13, 2026

How to use Speech 2.8 HD API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to minimax/speech-2.8-hd.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/tts",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "minimax/speech-2.8-hd",
      "text": "Hello from AI/ML API"
    },
)
print(r.json()["audio"])  # URL to the generated audio
const r = await fetch("https://api.aimlapi.com/v1/tts", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "minimax/speech-2.8-hd",
    "text": "Hello from AI/ML API"
  }),
});
const { audio } = await r.json();
console.log(audio); // URL to the generated audio
curl -X POST https://api.aimlapi.com/v1/tts \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"minimax/speech-2.8-hd","text":"Hello from AI/ML API"}'

# the JSON response contains "audio" — a URL to the generated speech

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Speech 2.8 HD API Pricing

TypePrice
Input
$130 / 1M tokens
Output
$130 / 1M tokens tokens

Speech 2.8 HD vs other models

ModelInputOutputContextBest for
Speech 2.8 HD
This page
$130 / 1M
$130 / 1M tokens
Speech synthesis
$130 / 1M
$0.13 / 1M tokens
Speech synthesis
$78 / 1M
$0.078 / 1M tokens
Speech synthesis
$13 / 1M
$0.013 / 1M tokens
Speech synthesis

Start building with Speech 2.8 HD

Get API Key
1000+ models, one API.