import requests r = requests.post( "https://api.aimlapi.com/v1/chat/completions", headers={"Authorization": "Bearer " + AIMLAPI_KEY}, json={ "model": "Qwen/Qwen2.5-7B-Instruct-Turbo", "messages": [ { "role": "user", "content": "Hello!" } ] }, ) print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", { method: "POST", headers: { Authorization: `Bearer ${process.env.AIMLAPI_KEY}`, "Content-Type": "application/json", }, body: JSON.stringify({ "model": "Qwen/Qwen2.5-7B-Instruct-Turbo", "messages": [ { "role": "user", "content": "Hello!" } ] }), }); console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \ -H "Authorization: Bearer $AIMLAPI_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"Qwen/Qwen2.5-7B-Instruct-Turbo","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Model | Input | Output | Context | Best for |
|---|---|---|---|---|
Qwen 2.5 7B Instruct Turbo This page | Coding + agents | |||
| Reasoning + agents | ||||
| Balanced coding + agents | ||||
| Long-context, multimodal & agentic workflows | ||||
| Reasoning + agents |
Qwen 2.5 7B Instruct Turbo has a 32,000 tokens context window and can return up to 31,000 tokens.
Qwen 2.5 7B Instruct Turbo takes text as input and returns text.
Use Qwen/Qwen2.5-7B-Instruct-Turbo as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.
Qwen 2.5 7B Instruct Turbo became available on September 30, 2025.
Qwen 2.5 7B Instruct Turbo is priced at input $0.39 / 1M tokens, output $0.39 / 1M tokens.
Yes, Qwen 2.5 7B Instruct Turbo can stream responses as they are generated.
Qwen 2.5 7B Instruct Turbo was built by Alibaba Cloud.
Yes, it supports function calling, tool use, and parallel tool calls.