Qwen3 VL 32B Instruct API

alibaba/qwen3-vl-32b-instruct
Qwen3 VL 32B Instruct can be seamlessly integrated into multimodal applications requiring precise image-text interaction.
Context
126K tokens
Input
$0.91 / 1M tokens
Output
$3.64 / 1M tokens
Released
Nov 11, 2025

How to use Qwen3 VL 32B Instruct API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to alibaba/qwen3-vl-32b-instruct.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "alibaba/qwen3-vl-32b-instruct",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "alibaba/qwen3-vl-32b-instruct",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"alibaba/qwen3-vl-32b-instruct","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Qwen3 VL 32B Instruct API Pricing

TypePrice
Input
$0.91 / 1M tokens
Output
$3.64 / 1M tokens

Qwen3 VL 32B Instruct Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
MMMU
76%
College-level multimodal understanding + reasoningSourceJuly 12, 2026
Intelligence
8.4
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026
Math
68.3
Composite score across standardised mathematics evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

Qwen3 VL 32B Instruct vs other models

ModelInputOutputContextBest for
$0.91 / 1M tokens
$3.64 / 1M tokens
126K tokens
Coding + agents
$6.5 / 1M tokens
$39 / 1M tokens
1M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$0.65 / 1M tokens
$3.9 / 1M tokens
1M tokens
Reasoning + agents

Frequently asked questions

Qwen3 VL 32B Instruct has a 126,000 tokens context window and can return up to 32,768 tokens.

Qwen3 VL 32B Instruct takes image, text as input and returns text.

Use alibaba/qwen3-vl-32b-instruct as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.

Qwen3 VL 32B Instruct became available on November 11, 2025.

Qwen3 VL 32B Instruct is priced at input $0.91 / 1M tokens, output $3.64 / 1M tokens.

Yes, Qwen3 VL 32B Instruct can stream responses as they are generated.

Yes, Qwen3 VL 32B Instruct accepts image input alongside text.

Qwen3 VL 32B Instruct was built by Alibaba Cloud.

Yes, it supports function calling, parallel tool calls, structured outputs, and web search.

Start building with Qwen3 VL 32B Instruct

Get API Key
1000+ models, one API.