Plain-text specification
Plain-text specification
GPT-5.5 API
GPT-5.5 is a large language model by OpenAI, available on the Venice API asopenai-gpt-55. Requests are anonymized, so the provider never sees your identity.GPT-5.5 is the latest frontier model in the GPT-5 series with a 1M+ context window, offering improved agentic and long context performance. It uses adaptive reasoning to dynamically allocate computation across tasks.GPT-5.5 API pricing
GPT-5.5 specifications
How to use the GPT-5.5 API
Send requests toPOST https://api.venice.ai/api/v1/responses with "model": "openai-gpt-55" and your API key.GPT-5.5 API FAQ
How much does the GPT-5.5 API cost?
37.50 per 1M output tokens, with cached input at $0.63 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the GPT-5.5 model ID?
Useopenai-gpt-55 as the model parameter.Is the GPT-5.5 API private?
GPT-5.5 is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.What is the context window of GPT-5.5?
1M tokens of context, with up to 128K output tokens per response.What does GPT-5.5 support?
GPT-5.5 supports function calling, structured outputs, reasoning, image input, web search and prompt caching. Reasoning effort is adjustable withreasoning_effort: none, low, medium, high and xhigh (default high).Which endpoint does the GPT-5.5 API use?
CallPOST /responses. /chat/completions is also supported. OpenAI reasoning models are designed around the Responses API: reasoning, tool calls and messages come back as typed output items. Venice’s /responses endpoint is in alpha.Related models
- GPT-5.4 API: 18.80 output per 1M tokens
- GPT-5.2 API: 17.50 output per 1M tokens
- Claude Opus 4.7 API: 30.00 output per 1M tokens
- Claude Opus 4.8 API: 30.00 output per 1M tokens
- Claude Opus 4.6 API: 30.00 output per 1M tokens
- Claude Opus 5 API: 30.00 output per 1M tokens