Skip to main content

Qwen 3.5 397B API

Qwen 3.5 397B is a large language model by Alibaba Qwen, available on the Venice API as qwen3-5-397b-a17b. Requests are anonymized, so the provider never sees your identity.Qwen 3.5 is Alibaba flagship reasoning model featuring a 397B parameter Mixture-of-Experts architecture with 17B active parameters. It excels at complex reasoning, coding, and general knowledge tasks.

Qwen 3.5 397B API pricing

Qwen 3.5 397B specifications

How to use the Qwen 3.5 397B API

Send requests to POST https://api.venice.ai/api/v1/chat/completions with "model": "qwen3-5-397b-a17b" and your API key.

Qwen 3.5 397B API FAQ

How much does the Qwen 3.5 397B API cost?

0.75per1Minputtokensand0.75 per 1M input tokens and 4.50 per 1M output tokens. Prices are in USD and can be paid in DIEM at parity.

What is the Qwen 3.5 397B model ID?

Use qwen3-5-397b-a17b as the model parameter.

Is the Qwen 3.5 397B API private?

Qwen 3.5 397B is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.

What is the context window of Qwen 3.5 397B?

125K tokens of context, with up to 32K output tokens per response.

What does Qwen 3.5 397B support?

Qwen 3.5 397B supports function calling, structured outputs, reasoning, image input and web search. Reasoning effort is adjustable with reasoning_effort: none, low, medium and high (default low).

Which endpoint does the Qwen 3.5 397B API use?

Call POST /chat/completions. /responses (Alpha) is also supported.

Related models

  • Qwen 3.8 27B API: 0.45input/0.45 input / 3.20 output per 1M tokens
  • Qwen 3.6 27B API: 0.33input/0.33 input / 3.25 output per 1M tokens
  • Qwen 2.5 7B API: 0.05input/0.05 input / 0.13 output per 1M tokens
  • Qwen 3.5 9B API: 0.10input/0.10 input / 0.15 output per 1M tokens
  • Grok 4.20 API: 1.42input/1.42 input / 2.83 output per 1M tokens
  • Grok 4.20 Multi-Agent API: 1.42input/1.42 input / 2.83 output per 1M tokens
  • GLM 5 API: 1.00input/1.00 input / 3.20 output per 1M tokens
  • Grok 4.3 API: 1.42input/1.42 input / 2.83 output per 1M tokens