Skip to main content

GPT-5.6 Luna Pro API

GPT-5.6 Luna Pro is a large language model by OpenAI, available on the Venice API as openai-gpt-56-luna-pro. Requests are anonymized, so the provider never sees your identity.GPT-5.6 Luna Pro is the same underlying model as GPT-5.6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

GPT-5.6 Luna Pro API pricing

GPT-5.6 Luna Pro specifications

How to use the GPT-5.6 Luna Pro API

Send requests to POST https://api.venice.ai/api/v1/responses with "model": "openai-gpt-56-luna-pro" and your API key.

GPT-5.6 Luna Pro API FAQ

How much does the GPT-5.6 Luna Pro API cost?

0.25per1Minputtokensand0.25 per 1M input tokens and 1.50 per 1M output tokens, with cached input at $0.025 per 1M. Prices are in USD and can be paid in DIEM at parity.

What is the GPT-5.6 Luna Pro model ID?

Use openai-gpt-56-luna-pro as the model parameter.

Is the GPT-5.6 Luna Pro API private?

GPT-5.6 Luna Pro is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.

What is the context window of GPT-5.6 Luna Pro?

1M tokens of context, with up to 125K output tokens per response.

What does GPT-5.6 Luna Pro support?

GPT-5.6 Luna Pro supports function calling, structured outputs, reasoning, image input, web search and prompt caching. Reasoning effort is adjustable with reasoning_effort: low, medium, high, xhigh and max (default medium).

Which endpoint does the GPT-5.6 Luna Pro API use?

Call POST /responses. /chat/completions is also supported. Pro and Codex models are Responses-native upstream. Use /responses for typed reasoning, tool-call and message items; /chat/completions stays fully supported.

Related models