Skip to main content

GPT-6 Sol API

GPT-6 Sol is a large language model by OpenAI, available on the Venice API as openai-gpt-6-sol. Requests are anonymized, so the provider never sees your identity.GPT-6 Sol is built for complex coding and agentic workflows. It is OpenAI’s GPT-6 coding and agent model, with a 1.05M token context window (922K input, 128K output), support for text and image inputs, and reasoning effort from none through max.

GPT-6 Sol API pricing

GPT-6 Sol specifications

How to use the GPT-6 Sol API

Send requests to POST https://api.venice.ai/api/v1/responses with "model": "openai-gpt-6-sol" and your API key.

GPT-6 Sol API FAQ

How much does the GPT-6 Sol API cost?

2.50per1Minputtokensand2.50 per 1M input tokens and 12.50 per 1M output tokens, with cached input at $0.25 per 1M. Prices are in USD and can be paid in DIEM at parity.

What is the GPT-6 Sol model ID?

Use openai-gpt-6-sol as the model parameter.

Is the GPT-6 Sol API private?

GPT-6 Sol is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.

What is the context window of GPT-6 Sol?

1.05M tokens of context, with up to 125K output tokens per response.

What does GPT-6 Sol support?

GPT-6 Sol supports function calling, structured outputs, reasoning, image input, web search and prompt caching. Reasoning effort is adjustable with reasoning_effort: none, low, medium, high, xhigh and max (default high).

Which endpoint does the GPT-6 Sol API use?

Call POST /responses. /chat/completions is also supported. OpenAI reasoning models are designed around the Responses API: reasoning, tool calls and messages come back as typed output items. Venice’s /responses endpoint is in alpha.

Related models

  • GPT-5.6 Sol API: 5.00input/5.00 input / 25.00 output per 1M tokens
  • Aion 3.5 API: 3.75input/3.75 input / 7.50 output per 1M tokens
  • Aion 3.0 API: 3.75input/3.75 input / 7.50 output per 1M tokens
  • Claude Sonnet 5 API: 3.00input/3.00 input / 15.00 output per 1M tokens
  • Qwen 3.8 2.4T API: 2.50input/2.50 input / 7.50 output per 1M tokens