import requests r = requests.post( "https://api.aimlapi.com/v1/chat/completions", headers={"Authorization": "Bearer " + AIMLAPI_KEY}, json={ "model": "openai/gpt-5.6-luna-pro", "messages": [ { "role": "user", "content": "Hello!" } ] }, ) print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", { method: "POST", headers: { Authorization: `Bearer ${process.env.AIMLAPI_KEY}`, "Content-Type": "application/json", }, body: JSON.stringify({ "model": "openai/gpt-5.6-luna-pro", "messages": [ { "role": "user", "content": "Hello!" } ] }), }); console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \ -H "Authorization: Bearer $AIMLAPI_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"openai/gpt-5.6-luna-pro","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Cached input |
GPT-5.6 Luna Pro has documented cache-write billing at 1.25× the uncached input token rate, and prompts with more than 272K input tokens are billed at 2× input and 1.5× output for the full request.
| Model | Input | Output | Context | Best for |
|---|---|---|---|---|
GPT-5.6 Luna Pro This page | Agentic workflows and structured output |
GPT-5.6 Luna Pro has a 1,050,000 tokens context window and can return up to 128,000 tokens.
GPT-5.6 Luna Pro takes image, text as input and returns text.
Use openai/gpt-5.6-luna-pro as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.
GPT-5.6 Luna Pro became available on July 9, 2026.
GPT-5.6 Luna Pro is priced at input $0.27508 / 1M tokens, output $1.65048 / 1M tokens, cached input $0.027508 / 1M tokens.
Yes, GPT-5.6 Luna Pro can stream responses as they are generated.
Yes, GPT-5.6 Luna Pro accepts image input alongside text.
GPT-5.6 Luna Pro was built by OpenAI.
Yes, it supports tools including parallel tool calls.