Skip to main content

Qwen 3 Coder 480B Turbo API

Qwen 3 Coder 480B Turbo is a large language model by Alibaba Qwen, available on the Venice API as qwen3-coder-480b-a35b-instruct-turbo. It runs privately, with zero data retention.Turbo variant of Qwen3 Coder 480B, optimized for faster inference on code tasks.

Qwen 3 Coder 480B Turbo API pricing

Qwen 3 Coder 480B Turbo specifications

How to use the Qwen 3 Coder 480B Turbo API

Send requests to POST https://api.venice.ai/api/v1/chat/completions with "model": "qwen3-coder-480b-a35b-instruct-turbo" and your API key.

Qwen 3 Coder 480B Turbo API FAQ

How much does the Qwen 3 Coder 480B Turbo API cost?

0.35per1Minputtokensand0.35 per 1M input tokens and 1.50 per 1M output tokens, with cached input at $0.04 per 1M. Prices are in USD and can be paid in DIEM at parity.

What is the Qwen 3 Coder 480B Turbo model ID?

Use qwen3-coder-480b-a35b-instruct-turbo as the model parameter.

Is the Qwen 3 Coder 480B Turbo API private?

Qwen 3 Coder 480B Turbo is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.

What is the context window of Qwen 3 Coder 480B Turbo?

250K tokens of context, with up to 64K output tokens per response.

What does Qwen 3 Coder 480B Turbo support?

Qwen 3 Coder 480B Turbo supports function calling, structured outputs, web search and prompt caching.

Which endpoint does the Qwen 3 Coder 480B Turbo API use?

Call POST /chat/completions. /responses (Alpha) is also supported.

Related models