Qwen3.5 Omni Flash API

alibaba/qwen3.5-omni-flash
Qwen3.5 Omni Flash scores well above the median for non-reasoning models in its price tier on the Artificial Analysis Intelligence Index (26 vs. a median of 15), generates output at 164 tokens per second, and supports every major input modality out of the box. It's competitively priced but runs verbose, producing roughly 2.5x more output tokens than average, which affects billing on output-heavy tasks.
Context
256K tokens
Input
$0.52 / 1M tokens
Output
$2.86 / 1M tokens
Released
Apr 24, 2026

How to use Qwen3.5 Omni Flash API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to alibaba/qwen3.5-omni-flash.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "alibaba/qwen3.5-omni-flash",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "alibaba/qwen3.5-omni-flash",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"alibaba/qwen3.5-omni-flash","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Qwen3.5 Omni Flash API Pricing

TypePrice
Input
$0.52 / 1M tokens
Output
$2.86 / 1M tokens

Qwen3.5 Omni Flash Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
Intelligence
12.5
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

Qwen3.5 Omni Flash vs other models

ModelInputOutputContextBest for
$0.52 / 1M tokens
$2.86 / 1M tokens
256K tokens
Reasoning + agents
$6.5 / 1M tokens
$39 / 1M tokens
1.05M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$0.65 / 1M tokens
$3.9 / 1M tokens
1.05M tokens
Reasoning + agents

Frequently asked questions

Yes, reasoning is listed among its features and capabilities.

Start building with Qwen3.5 Omni Flash

Get API Key
1000+ models, one API.