GPT-5.6 Luna Pro API

openai/gpt-5.6-luna-pro
OpenAI's extended-reasoning configuration of GPT-5.6 Luna, trading speed for deeper reasoning on complex requests, with a 1M-token context window.
Context
1.1M tokens
Input
$0.27508 / 1M tokens
Output
$1.65048 / 1M tokens
Released
Jul 9, 2026

How to use GPT-5.6 Luna Pro API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to openai/gpt-5.6-luna-pro.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "openai/gpt-5.6-luna-pro",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "openai/gpt-5.6-luna-pro",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"openai/gpt-5.6-luna-pro","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

GPT-5.6 Luna Pro API Pricing

TypePrice
Input
$0.27508 / 1M tokens
Output
$1.65048 / 1M tokens
Cached input
$0.027508 / 1M tokens

GPT-5.6 Luna Pro has documented cache-write billing at 1.25× the uncached input token rate, and prompts with more than 272K input tokens are billed at 2× input and 1.5× output for the full request.

GPT-5.6 Luna Pro vs other models

ModelInputOutputContextBest for
$0.27508 / 1M tokens
$1.65048 / 1M tokens
1.1M tokens
Agentic workflows and structured output
$0.0975 / 1M tokens
$0.325 / 1M tokens
1M tokens
Coding + agentic reasoning
$0.82524 / 1M tokens
$8.2524 / 1M tokens
1.04M tokens
Reasoning + agents
$0.41262 / 1M tokens
$3.4385 / 1M tokens
1M tokens
Reasoning + agents

Frequently asked questions

GPT-5.6 Luna Pro has a 1,050,000 tokens context window and can return up to 128,000 tokens.

GPT-5.6 Luna Pro takes image, text as input and returns text.

Use openai/gpt-5.6-luna-pro as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.

GPT-5.6 Luna Pro became available on July 9, 2026.

GPT-5.6 Luna Pro is priced at input $0.27508 / 1M tokens, output $1.65048 / 1M tokens, cached input $0.027508 / 1M tokens.

Yes, GPT-5.6 Luna Pro can stream responses as they are generated.

Yes, GPT-5.6 Luna Pro accepts image input alongside text.

GPT-5.6 Luna Pro was built by OpenAI.

Yes, it supports tools including parallel tool calls.

Start building with GPT-5.6 Luna Pro

Get API Key
1000+ models, one API.