Plain-text specification
Plain-text specification
Qwen3 VL 235B API
Qwen3 VL 235B is a large language model by Alibaba Qwen, available on the Venice API asqwen3-vl-235b-a22b. It runs privately, with zero data retention.Qwen3-VL 235B vision-language model with MoE architecture. The most powerful VL model in the Qwen series with superior visual perception, OCR, and multimodal reasoning.Qwen3 VL 235B API pricing
Qwen3 VL 235B specifications
How to use the Qwen3 VL 235B API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "qwen3-vl-235b-a22b" and your API key.Qwen3 VL 235B API FAQ
How much does the Qwen3 VL 235B API cost?
1.90 per 1M output tokens, with cached input at $0.10 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the Qwen3 VL 235B model ID?
Useqwen3-vl-235b-a22b as the model parameter.Is the Qwen3 VL 235B API private?
Qwen3 VL 235B is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.What is the context window of Qwen3 VL 235B?
125K tokens of context, with up to 16K output tokens per response.What does Qwen3 VL 235B support?
Qwen3 VL 235B supports function calling, structured outputs, image input, web search and prompt caching.Which endpoint does the Qwen3 VL 235B API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- MiniMax M2.7 API: 1.50 output per 1M tokens
- Mercury 2 API: 0.94 output per 1M tokens
- GPT-5.6 Luna API: 1.50 output per 1M tokens
- GPT-5.6 Luna Pro API: 1.50 output per 1M tokens