Plain-text specification
Plain-text specification
Qwen 3 Next 80b API
Qwen 3 Next 80b is a large language model by Alibaba Qwen, available on the Venice API asqwen3-next-80b. It runs privately, with zero data retention.Optimized for speed and efficiency.Qwen 3 Next 80b API pricing
Qwen 3 Next 80b specifications
How to use the Qwen 3 Next 80b API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "qwen3-next-80b" and your API key.Qwen 3 Next 80b API FAQ
How much does the Qwen 3 Next 80b API cost?
1.90 per 1M output tokens. Prices are in USD and can be paid in DIEM at parity.What is the Qwen 3 Next 80b model ID?
Useqwen3-next-80b as the model parameter.Is the Qwen 3 Next 80b API private?
Qwen 3 Next 80b is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.What is the context window of Qwen 3 Next 80b?
250K tokens of context, with up to 16K output tokens per response.What does Qwen 3 Next 80b support?
Qwen 3 Next 80b supports function calling, structured outputs and web search.Which endpoint does the Qwen 3 Next 80b API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- Llama 3.3 70B API: 2.80 output per 1M tokens
- MiniMax M2.7 API: 1.50 output per 1M tokens
- GLM 4.6 API: 1.75 output per 1M tokens
- Venice Role Play Uncensored API: 2.00 output per 1M tokens