GLM 5V Turbo API

z-ai/glm-5v-turbo
GLM 5V Turbo on AIMLAPI.
Context
203K tokens
Input
$1.56 / 1M tokens
Output
$5.2 / 1M tokens

How to use GLM 5V Turbo API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to z-ai/glm-5v-turbo.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "z-ai/glm-5v-turbo",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "z-ai/glm-5v-turbo",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-5v-turbo","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

GLM 5V Turbo API Pricing

TypePrice
Input
$1.56 / 1M tokens
Output
$5.2 / 1M tokens
Cached input
$0.312 / 1M tokens

GLM 5V Turbo Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
Intelligence
23.5
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

GLM 5V Turbo vs other models

ModelInputOutputContextBest for
GLM 5V Turbo
This page
$1.56 / 1M tokens
$5.2 / 1M tokens
203K tokens
Reasoning + agents
$6.5 / 1M tokens
$39 / 1M tokens
1.05M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$0.65 / 1M tokens
$3.9 / 1M tokens
1.05M tokens
Reasoning + agents

Frequently asked questions

GLM 5V Turbo has a 202,752 tokens context window and can return up to 131,072 tokens.

GLM 5V Turbo takes image, text as input and returns text.

Use z-ai/glm-5v-turbo as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.

GLM 5V Turbo is priced at input $1.56 / 1M tokens, output $5.2 / 1M tokens, cached input $0.312 / 1M tokens.

Yes, GLM 5V Turbo can stream responses as they are generated.

Yes, GLM 5V Turbo accepts image input alongside text.

GLM 5V Turbo was built by Zhipu AI.

Yes, it supports function calling, tool use, and parallel tool calls.

Send requests to https://api.aimlapi.com/v1/chat/completions using the model id z-ai/glm-5v-turbo.

Start building with GLM 5V Turbo

Get API Key
1000+ models, one API.