Plain-text specification
Plain-text specification
Qwen 3 Coder 480B Turbo API
Qwen 3 Coder 480B Turbo is a large language model by Alibaba Qwen, available on the Venice API asqwen3-coder-480b-a35b-instruct-turbo. It runs privately, with zero data retention.Turbo variant of Qwen3 Coder 480B, optimized for faster inference on code tasks.Qwen 3 Coder 480B Turbo API pricing
Qwen 3 Coder 480B Turbo specifications
How to use the Qwen 3 Coder 480B Turbo API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "qwen3-coder-480b-a35b-instruct-turbo" and your API key.Qwen 3 Coder 480B Turbo API FAQ
How much does the Qwen 3 Coder 480B Turbo API cost?
1.50 per 1M output tokens, with cached input at $0.04 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the Qwen 3 Coder 480B Turbo model ID?
Useqwen3-coder-480b-a35b-instruct-turbo as the model parameter.Is the Qwen 3 Coder 480B Turbo API private?
Qwen 3 Coder 480B Turbo is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.What is the context window of Qwen 3 Coder 480B Turbo?
250K tokens of context, with up to 64K output tokens per response.What does Qwen 3 Coder 480B Turbo support?
Qwen 3 Coder 480B Turbo supports function calling, structured outputs, web search and prompt caching.Which endpoint does the Qwen 3 Coder 480B Turbo API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- MiniMax M2.7 API: 1.50 output per 1M tokens
- DeepSeek V4.1 Flash API: 1.50 output per 1M tokens
- Mercury 2 API: 0.94 output per 1M tokens
- GPT-5.6 Luna API: 1.50 output per 1M tokens