Seed Audio 1.0 API

bytedance/seed-audio-1-0
Seed Audio 1.0 is ByteDance's non-streaming audio generation model, producing speech, sound effects, and reference-guided audio with precise timbre control.
Output
$0.00325 / sec
Released
Jun 29, 2026

How to use Seed Audio 1.0 API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to bytedance/seed-audio-1-0.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/tts",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "bytedance/seed-audio-1-0",
      "text": "Hello from AI/ML API"
    },
)
print(r.json()["audio"])  # URL to the generated audio
const r = await fetch("https://api.aimlapi.com/v1/tts", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "bytedance/seed-audio-1-0",
    "text": "Hello from AI/ML API"
  }),
});
const { audio } = await r.json();
console.log(audio); // URL to the generated audio
curl -X POST https://api.aimlapi.com/v1/tts \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"bytedance/seed-audio-1-0","text":"Hello from AI/ML API"}'

# the JSON response contains "audio" — a URL to the generated speech

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Seed Audio 1.0 API Pricing

TypePrice
Output
$0.00325 / sec

Seed Audio 1.0 vs other models

ModelInputOutputContextBest for
$0.00325 / sec
Speech synthesis
$130 / 1M characters
$130 / 1M
Speech synthesis
$130 / 1M characters
$0.13 / 1M tokens
Speech synthesis
$78 / 1M characters
$0.078 / 1M tokens
Speech synthesis
$13 / 1M characters
$0.013 / 1M tokens
Speech synthesis

Frequently asked questions

Seed Audio 1.0 takes text as input and returns audio.

Seed Audio 1.0 became available on June 29, 2026.

Seed Audio 1.0 is priced at $0.00325 / sec.

Seed Audio 1.0 is billed per generation — a fixed charge per output rather than by prompt length.

Seed Audio 1.0 was built by ByteDance.

Send a request to https://api.aimlapi.com/v1/tts with bytedance/seed-audio-1-0 as the model id.

Yes. Seed Audio 1.0 is served through AI/ML API, so the same key and endpoint format used for other models applies.

No. Besides converting text to speech, it also generates sound effects, offers timbre control, and creates audio guided by a reference input.

No, it is a non-streaming model, meaning the full audio output is produced after processing rather than streamed as it's generated.

Start building with Seed Audio 1.0

Get API Key
1000+ models, one API.