Stable Audio API

stable-audio
Stable Audio generates high-quality audio from text prompts with innovative features like audio transformation and extensive creative control.
Output
$0.000156 / step (variable)

How to use Stable Audio API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to stable-audio.
import requests, time

headers = {"Authorization": "Bearer " + AIMLAPI_KEY}
job = requests.post(
    "https://api.aimlapi.com/v2/generate/audio",
    headers=headers,
    json={
      "model": "stable-audio",
      "prompt": "Upbeat lofi background music"
    },
).json()
gid = job["generation_id"]

while True:
    res = requests.get(f"https://api.aimlapi.com/v2/generate/audio?generation_id={gid}", headers=headers).json()
    if res.get("status") in ("completed", "error", "failed"):
        break
    time.sleep(3)
print(res)
const headers = {
  Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
  "Content-Type": "application/json",
};
const job = await (await fetch("https://api.aimlapi.com/v2/generate/audio", {
  method: "POST",
  headers,
  body: JSON.stringify({
    "model": "stable-audio",
    "prompt": "Upbeat lofi background music"
  }),
})).json();

let res;
do {
  await new Promise((r) => setTimeout(r, 3000));
  res = await (await fetch(`https://api.aimlapi.com/v2/generate/audio?generation_id=${job.generation_id}`, { headers })).json();
} while (!["completed", "error", "failed"].includes(res.status));
console.log(res);
# submit the job — the response contains "generation_id"
curl -X POST https://api.aimlapi.com/v2/generate/audio \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"stable-audio","prompt":"Upbeat lofi background music"}'

# then poll for the result until it is ready
curl "https://api.aimlapi.com/v2/generate/audio?generation_id={generation_id}" -H "Authorization: Bearer $AIMLAPI_KEY"

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Stable Audio API Pricing

TypePrice
Output
$0.000156 / step (variable)

Stable Audio vs other models

ModelInputOutputContextBest for
Stable Audio
This page
$0.000156 / step (variable)
Audio generation
$130 / 1M characters
$130 / 1M
Speech synthesis
$130 / 1M characters
$0.13 / 1M tokens
Speech synthesis
$78 / 1M characters
$0.078 / 1M tokens
Speech synthesis
$13 / 1M characters
$0.013 / 1M tokens
Speech synthesis

Frequently asked questions

Stable Audio takes audio, text as input and returns audio.

Stable Audio is priced at output $0.000156 / step (variable).

Stable Audio was built by Stability AI.

Send a request to v2/generate/audio with stable-audio as the model id.

Yes. Stable Audio is served through AI/ML API, so the same key and endpoint format used for other models applies.

Stable Audio includes audio transformation features and gives users extensive creative control over the generated output.

Start building with Stable Audio

Get API Key
1000+ models, one API.