Plain-text specification
Plain-text specification
Qwen 3.5 397B API
Qwen 3.5 397B is a large language model by Alibaba Qwen, available on the Venice API asqwen3-5-397b-a17b. Requests are anonymized, so the provider never sees your identity.Qwen 3.5 is Alibaba flagship reasoning model featuring a 397B parameter Mixture-of-Experts architecture with 17B active parameters. It excels at complex reasoning, coding, and general knowledge tasks.Qwen 3.5 397B API pricing
Qwen 3.5 397B specifications
How to use the Qwen 3.5 397B API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "qwen3-5-397b-a17b" and your API key.Qwen 3.5 397B API FAQ
How much does the Qwen 3.5 397B API cost?
4.50 per 1M output tokens. Prices are in USD and can be paid in DIEM at parity.What is the Qwen 3.5 397B model ID?
Useqwen3-5-397b-a17b as the model parameter.Is the Qwen 3.5 397B API private?
Qwen 3.5 397B is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.What is the context window of Qwen 3.5 397B?
125K tokens of context, with up to 32K output tokens per response.What does Qwen 3.5 397B support?
Qwen 3.5 397B supports function calling, structured outputs, reasoning, image input and web search. Reasoning effort is adjustable withreasoning_effort: none, low, medium and high (default low).Which endpoint does the Qwen 3.5 397B API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- Qwen 3.8 27B API: 3.20 output per 1M tokens
- Qwen 3.6 27B API: 3.25 output per 1M tokens
- Qwen 2.5 7B API: 0.13 output per 1M tokens
- Qwen 3.5 9B API: 0.15 output per 1M tokens
- Grok 4.20 API: 2.83 output per 1M tokens
- Grok 4.20 Multi-Agent API: 2.83 output per 1M tokens
- GLM 5 API: 3.20 output per 1M tokens
- Grok 4.3 API: 2.83 output per 1M tokens