Plain-text specification
Plain-text specification
Qwen3 VL 30B A3B API
Qwen3 VL 30B A3B is a large language model by Alibaba Qwen, available on the Venice API ase2ee-qwen3-vl-30b-a3b-p. It is end-to-end encrypted.Qwen3 VL 30B A3B running in a Trusted Execution Environment (TEE). A multimodal model unifying text generation with visual understanding for images and videos, with hardware attestation evidence available for independent verification.Qwen3 VL 30B A3B API pricing
Qwen3 VL 30B A3B specifications
How to use the Qwen3 VL 30B A3B API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "e2ee-qwen3-vl-30b-a3b-p" and your API key.Qwen3 VL 30B A3B API FAQ
How much does the Qwen3 VL 30B A3B API cost?
0.90 per 1M output tokens. Prices are in USD and can be paid in DIEM at parity.What is the Qwen3 VL 30B A3B model ID?
Usee2ee-qwen3-vl-30b-a3b-p as the model parameter.Is the Qwen3 VL 30B A3B API private?
Qwen3 VL 30B A3B is end-to-end encrypted: prompts are encrypted on your device and decrypted only inside an attested hardware enclave, so neither Venice nor the GPU provider can read them.What is the context window of Qwen3 VL 30B A3B?
125K tokens of context, with up to 4K output tokens per response.What does Qwen3 VL 30B A3B support?
Qwen3 VL 30B A3B supports function calling, image input and web search.Which endpoint does the Qwen3 VL 30B A3B API use?
CallPOST /chat/completions.Related models
- MiniMax M2.5 API: 0.95 output per 1M tokens
- Venice Uncensored 1.2 API: 0.90 output per 1M tokens
- Mercury 2 API: 0.94 output per 1M tokens
- Gemma 4 26B A4B Uncensored API: 0.88 output per 1M tokens