import requests
r = requests.post(
"https://api.aimlapi.com/v1/chat/completions",
headers={"Authorization": "Bearer " + AIMLAPI_KEY},
json={
"model": "openai/gpt-5.6-luna",
"messages": [
{
"role": "user",
"content": "Hello!"
}
]
},
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
method: "POST",
headers: {
Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
"model": "openai/gpt-5.6-luna",
"messages": [
{
"role": "user",
"content": "Hello!"
}
]
}),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
-H "Authorization: Bearer $AIMLAPI_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"openai/gpt-5.6-luna","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Cached input |
Prompt caching: cache reads get a 90% discount, cache writes are billed at 1.25x input rate. Free credits on signup.
| Benchmark | Score | What it measures | Source | Retrieved |
|---|---|---|---|---|
| Intelligence | 37.5 | Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
| Coding | 71.4 | Composite score across standardised coding evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
| Model | Input | Output | Context | Best for |
|---|---|---|---|---|
GPT-5.6 Luna This page | Fast, high-volume tasks | |||
| Reasoning + agents | ||||
| Reasoning + agents | ||||
| Chat + assistants | ||||
| Reasoning + agents |
GPT-5.6 Luna has a 1,050,000 tokens context window and can return up to 128,000 tokens.
GPT-5.6 Luna takes text, image, file in as input and returns text out.
Use openai/gpt-5.6-luna as the model id.
GPT-5.6 Luna became available on July 9, 2026.
GPT-5.6 Luna is priced at input $0.26 / 1M tokens, output $1.56 / 1M tokens, cached input $0.026 / 1M tokens.
Yes, GPT-5.6 Luna can stream responses as they are generated.
Yes, GPT-5.6 Luna accepts image input alongside text.
GPT-5.6 Luna was built by OpenAI.
It is a chat and code model with reasoning, vision, and tool-calling capabilities, making it suitable for conversational and coding tasks that require structured output.