Qwen 2.5 7B Instruct Turbo API

Qwen/Qwen2.5-7B-Instruct-Turbo
Qwen 2.5 7B Instruct Turbo excels in coding and instruction following.
Context
32K tokens
Input
$0.39 / 1M tokens
Output
$0.39 / 1M tokens
Released
Sep 30, 2025

How to use Qwen 2.5 7B Instruct Turbo API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to Qwen/Qwen2.5-7B-Instruct-Turbo.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "Qwen/Qwen2.5-7B-Instruct-Turbo",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "Qwen/Qwen2.5-7B-Instruct-Turbo",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"Qwen/Qwen2.5-7B-Instruct-Turbo","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Qwen 2.5 7B Instruct Turbo API Pricing

TypePrice
Input
$0.39 / 1M tokens
Output
$0.39 / 1M tokens

Qwen 2.5 7B Instruct Turbo vs other models

ModelInputOutputContextBest for
$0.39 / 1M tokens
$0.39 / 1M tokens
32K tokens
Coding + agents
$6.5 / 1M tokens
$39 / 1M tokens
1.05M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$0.65 / 1M tokens
$3.9 / 1M tokens
1.05M tokens
Reasoning + agents

Frequently asked questions

Qwen 2.5 7B Instruct Turbo has a 32,000 tokens context window and can return up to 31,000 tokens.

Qwen 2.5 7B Instruct Turbo takes text as input and returns text.

Use Qwen/Qwen2.5-7B-Instruct-Turbo as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.

Qwen 2.5 7B Instruct Turbo became available on September 30, 2025.

Qwen 2.5 7B Instruct Turbo is priced at input $0.39 / 1M tokens, output $0.39 / 1M tokens.

Yes, Qwen 2.5 7B Instruct Turbo can stream responses as they are generated.

Qwen 2.5 7B Instruct Turbo was built by Alibaba Cloud.

Yes, it supports function calling, tool use, and parallel tool calls.

Start building with Qwen 2.5 7B Instruct Turbo

Get API Key
1000+ models, one API.