import requests r = requests.post( "https://api.aimlapi.com/v1/chat/completions", headers={"Authorization": "Bearer " + AIMLAPI_KEY}, json={ "model": "z-ai/glm-5v-turbo", "messages": [ { "role": "user", "content": "Hello!" } ] }, ) print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", { method: "POST", headers: { Authorization: `Bearer ${process.env.AIMLAPI_KEY}`, "Content-Type": "application/json", }, body: JSON.stringify({ "model": "z-ai/glm-5v-turbo", "messages": [ { "role": "user", "content": "Hello!" } ] }), }); console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \ -H "Authorization: Bearer $AIMLAPI_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"z-ai/glm-5v-turbo","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Cached input |
| Benchmark | Score | What it measures | Source | Retrieved |
|---|---|---|---|---|
| Intelligence | 23.5 | Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
| Model | Input | Output | Context | Best for |
|---|---|---|---|---|
GLM 5V Turbo This page | Reasoning + agents | |||
| Reasoning + agents | ||||
| Balanced coding + agents | ||||
| Long-context, multimodal & agentic workflows | ||||
| Reasoning + agents |
GLM 5V Turbo has a 202,752 tokens context window and can return up to 131,072 tokens.
GLM 5V Turbo takes image, text as input and returns text.
Use z-ai/glm-5v-turbo as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.
GLM 5V Turbo is priced at input $1.56 / 1M tokens, output $5.2 / 1M tokens, cached input $0.312 / 1M tokens.
Yes, GLM 5V Turbo can stream responses as they are generated.
Yes, GLM 5V Turbo accepts image input alongside text.
GLM 5V Turbo was built by Zhipu AI.
Yes, it supports function calling, tool use, and parallel tool calls.
Send requests to https://api.aimlapi.com/v1/chat/completions using the model id z-ai/glm-5v-turbo.