Skip to main content

Qwen 3.8 27B API

Qwen 3.8 27B is a large language model by Alibaba Qwen, available on the Venice API as qwen-3-8-27b. Private and end-to-end encrypted variants are available.Qwen 3.8 27B is a native vision-language dense model with 27B parameters. It improves coding, professional work, research, and long-horizon agentic tasks, with flexible thinking control and image and video understanding. It supports a native 262K-token context window.

Qwen 3.8 27B API pricing

Qwen 3.8 27B specifications

How to use the Qwen 3.8 27B API

Send requests to POST https://api.venice.ai/api/v1/chat/completions with "model": "qwen-3-8-27b" and your API key.

Qwen 3.8 27B API FAQ

How much does the Qwen 3.8 27B API cost?

0.45per1Minputtokensand0.45 per 1M input tokens and 3.20 per 1M output tokens. The E2EE variant costs 0.47inputand0.47 input and 3.53 output. Prices are in USD and can be paid in DIEM at parity.

What is the Qwen 3.8 27B model ID?

Use qwen-3-8-27b as the model parameter. Other variants: e2ee-qwen3-8-27b (E2EE).

Is the Qwen 3.8 27B API private?

The Standard variant is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training. The E2EE variant is end-to-end encrypted: prompts are encrypted on your device and decrypted only inside an attested hardware enclave, so neither Venice nor the GPU provider can read them.

What is the context window of Qwen 3.8 27B?

256K tokens of context, with up to 64K output tokens per response.

What does Qwen 3.8 27B support?

Qwen 3.8 27B supports function calling, structured outputs, reasoning, image input and web search. Reasoning effort is adjustable with reasoning_effort: none, low, medium and xhigh (default xhigh).

Which endpoint does the Qwen 3.8 27B API use?

Call POST /chat/completions. /responses (Alpha) is also supported.

Related models