import requests r = requests.post( "https://api.aimlapi.com/v1/chat/completions", headers={"Authorization": "Bearer " + AIMLAPI_KEY}, json={ "model": "alibaba/glm-5.2-fast-preview", "messages": [ { "role": "user", "content": "Hello!" } ] }, ) print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", { method: "POST", headers: { Authorization: `Bearer ${process.env.AIMLAPI_KEY}`, "Content-Type": "application/json", }, body: JSON.stringify({ "model": "alibaba/glm-5.2-fast-preview", "messages": [ { "role": "user", "content": "Hello!" } ] }), }); console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \ -H "Authorization: Bearer $AIMLAPI_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"alibaba/glm-5.2-fast-preview","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Cached input |
glm-5.2-fast-preview has no free quota on Alibaba Cloud’s Model Studio pricing page, and its input, output, and cached-input prices are listed separately for each deployment scope.
| Model | Input | Output | Context | Best for |
|---|---|---|---|---|
GLM 5.2 Fast Preview This page | Agentic workflows and structured output |
GLM 5.2 Fast Preview has a 1,000,000 tokens context window and can return up to 131,072 tokens.
GLM 5.2 Fast Preview takes text as input and returns text.
Use alibaba/glm-5.2-fast-preview as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.
GLM 5.2 Fast Preview became available on July 10, 2026.
GLM 5.2 Fast Preview is priced at input $4.55 / 1M tokens, output $14.3 / 1M tokens, cached input $0.91 / 1M tokens.
Yes, GLM 5.2 Fast Preview can stream responses as they are generated.
GLM 5.2 Fast Preview was built by Zhipu AI.
It supports reasoning, streaming, structured output, tools, and parallel tool calls.