import requests r = requests.post( "https://api.aimlapi.com/v1/chat/completions", headers={"Authorization": "Bearer " + AIMLAPI_KEY}, json={ "model": "openai/gpt-3.5-turbo", "messages": [ { "role": "user", "content": "Hello!" } ] }, ) print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", { method: "POST", headers: { Authorization: `Bearer ${process.env.AIMLAPI_KEY}`, "Content-Type": "application/json", }, body: JSON.stringify({ "model": "openai/gpt-3.5-turbo", "messages": [ { "role": "user", "content": "Hello!" } ] }), }); console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \ -H "Authorization: Bearer $AIMLAPI_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"openai/gpt-3.5-turbo","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Benchmark | Score | What it measures | Source | Retrieved |
|---|---|---|---|---|
| Intelligence | 5.5 | Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
| Coding | 10.7 | Composite score across standardised coding evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
GPT-3.5 Turbo has a 16,000 tokens context window.
GPT-3.5 Turbo takes text as input and returns text.
Use openai/gpt-3.5-turbo as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.
GPT-3.5 Turbo became available on August 7, 2025.
GPT-3.5 Turbo is priced at input $0.65 / 1M tokens, output $1.95 / 1M tokens.
Yes, GPT-3.5 Turbo can stream responses as they are generated.
GPT-3.5 Turbo was built by OpenAI.
No. Its listed capabilities are function calling, streaming, and structured outputs, not dedicated reasoning.
Yes. It is available at the endpoint https://api.aimlapi.com/v1/chat/completions.
Yes. It supports function calling, tools, and parallel tool calls.
Yes. It supports structured output generation.