OmniHuman v1.5 API

bytedance/omnihuman/v1.5
OmniHuman v1.5 is an advanced multimodal AI model designed to transform a single human image and an audio input into highly realistic video footage.
Output
$0.208 / sec

How to use OmniHuman v1.5 API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to bytedance/omnihuman/v1.5.
import requests, time

headers = {"Authorization": "Bearer " + AIMLAPI_KEY}
job = requests.post(
    "https://api.aimlapi.com/v2/video/generations",
    headers=headers,
    json={
      "model": "bytedance/omnihuman/v1.5",
      "prompt": "A serene timelapse of clouds over a mountain range"
    },
).json()
gid = job["id"]

while True:
    res = requests.get(f"https://api.aimlapi.com/v2/video/generations?generation_id={gid}", headers=headers).json()
    if res.get("status") in ("completed", "error"):
        break
    time.sleep(5)
print(res["video"]["url"])
const headers = {
  Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
  "Content-Type": "application/json",
};
const job = await (await fetch("https://api.aimlapi.com/v2/video/generations", {
  method: "POST",
  headers,
  body: JSON.stringify({
    "model": "bytedance/omnihuman/v1.5",
    "prompt": "A serene timelapse of clouds over a mountain range"
  }),
})).json();

let res;
do {
  await new Promise((r) => setTimeout(r, 5000));
  res = await (await fetch(`https://api.aimlapi.com/v2/video/generations?generation_id=${job.id}`, { headers })).json();
} while (!["completed", "error"].includes(res.status));
console.log(res.video.url);
# submit the generation — the response contains the job "id"
curl -X POST https://api.aimlapi.com/v2/video/generations \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"bytedance/omnihuman/v1.5","prompt":"A serene timelapse of clouds over a mountain range"}'

# then poll until status is "completed" — video.url holds the result
curl "https://api.aimlapi.com/v2/video/generations?generation_id={id}" -H "Authorization: Bearer $AIMLAPI_KEY"

OpenAI-compatible — swap the base URL and it works with your existing SDK.

OmniHuman v1.5 API Pricing

TypePrice
Output
$0.208 / sec

OmniHuman v1.5 vs other models

ModelInputOutputContextBest for
$0.208 / sec
Video generation
$0.26 / sec (variable)
Video generation
$0.13 / sec (variable)
Video generation
$1.95 / 1M tokens
$22.75 / 1M tokens
Video generation
$0.09243–$1.014 / sec (by resolution)
Cinematic video + native audio

Frequently asked questions

OmniHuman v1.5 takes image, text, video as input and returns video.

OmniHuman v1.5 is priced at $0.208 / sec.

OmniHuman v1.5 is billed per generation — a fixed charge per output rather than by prompt length.

Yes, OmniHuman v1.5 accepts image input as well as a text prompt.

OmniHuman v1.5 was built by ByteDance.

Send a request to v2/video/generations with bytedance/omnihuman/v1.5 as the model id.

Yes. OmniHuman v1.5 is served through AI/ML API, so the same key and endpoint format used for other models applies.

It generates videos of human figures with more natural motion and more accurate facial detail than prior versions.

Start building with OmniHuman v1.5

Get API Key
1000+ models, one API.