Processing...Please wait while we secure this action
OpenAI model

GPT-5.6 Luna

OpenAI: GPT-5.6 Luna is a OpenAI model listed in the NextModel catalog for image understanding, multimodal chat workloads. Its listed price is $0.029 / 1M tokens input and $0.174 / 1M tokens output per 1M tokens, with a 1.1M token context window.

OpenAIPlatform curatedCatalog
StreamingTool callingJSON modeVisionLong context
Input price$0.029 / 1M tokens
Output price$0.174 / 1M tokens
Context length1.1M tokens
Max output8.2k tokens
Input lengthInput / 1MOutput / 1MCache hit / 1M
— - 272k$0.029$0.174$0.0029
> 272k$0.058$0.26$0.0058

What is GPT-5.6 Luna in NextModel?

OpenAI: GPT-5.6 Luna is a OpenAI model listed in the NextModel catalog for image understanding, multimodal chat workloads. Its listed price is $0.029 / 1M tokens input and $0.174 / 1M tokens output per 1M tokens, with a 1.1M token context window.

Best use cases

  • image understanding
  • multimodal chat

OpenAI-compatible code example

Keep the OpenAI SDK style, set base_url to NextModel, and use the catalog model ID openai--gpt-5.6-luna.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="openai/gpt-5.6-luna",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Similar alternatives

OpenAICatalog

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

$0.289 / 1M tokensInput$1.16 / 1M tokensOutput1MContext
Best forimage understanding, multimodal chat
RoutingConfigured
StreamingTool callingJSON modeVision
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenAICatalog

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

$0.058 / 1M tokensInput$0.231 / 1M tokensOutput1MContext
Best forimage understanding, multimodal chat
RoutingConfigured
StreamingTool callingJSON modeVision
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenAICatalog

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

$0.014 / 1M tokensInput$0.058 / 1M tokensOutput1MContext
Best forimage understanding, multimodal chat
RoutingConfigured
StreamingTool callingJSON modeVision
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

OpenAI: GPT-5.6 Luna API questions