Plain-text specification
Plain-text specification
Qwen 3.7 Max API
Qwen 3.7 Max is a large language model by Alibaba Qwen, available on the Venice API asqwen-3-7-max. Requests are anonymized, so the provider never sees your identity.Qwen 3.7 Max is the largest model in the Qwen 3.7 series, with deep thinking, function calling, prompt caching, and multimodal input support for images and video. It excels at programming, office and productivity tasks, and long-running autonomous agent workflows.Qwen 3.7 Max API pricing
Qwen 3.7 Max specifications
How to use the Qwen 3.7 Max API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "qwen-3-7-max" and your API key.Qwen 3.7 Max API FAQ
How much does the Qwen 3.7 Max API cost?
8.05 per 1M output tokens, with cached input at $0.27 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the Qwen 3.7 Max model ID?
Useqwen-3-7-max as the model parameter.Is the Qwen 3.7 Max API private?
Qwen 3.7 Max is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.What is the context window of Qwen 3.7 Max?
1M tokens of context, with up to 64K output tokens per response.What does Qwen 3.7 Max support?
Qwen 3.7 Max supports function calling, reasoning, image input, web search and prompt caching.Which endpoint does the Qwen 3.7 Max API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- Qwen 3.8 Max API: 7.50 output per 1M tokens
- Gemini 3.5 Flash API: 9.45 output per 1M tokens
- Aion 3.0 API: 7.50 output per 1M tokens
- Grok 4.5 API: 6.80 output per 1M tokens
- Grok 4.6 API: 6.80 output per 1M tokens