Plain-text specification
Plain-text specification
Qwen 2.5 7B API
Qwen 2.5 7B is a large language model by Alibaba Qwen, available on the Venice API ase2ee-qwen-2-5-7b-p. It is end-to-end encrypted.Qwen 2.5 7B Instruct running in a Trusted Execution Environment (TEE). A compact model with strong coding, math, and multilingual capabilities supporting 29+ languages, with hardware attestation evidence available for independent verification.Qwen 2.5 7B API pricing
Qwen 2.5 7B specifications
How to use the Qwen 2.5 7B API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "e2ee-qwen-2-5-7b-p" and your API key.Qwen 2.5 7B API FAQ
How much does the Qwen 2.5 7B API cost?
0.13 per 1M output tokens. Prices are in USD and can be paid in DIEM at parity.What is the Qwen 2.5 7B model ID?
Usee2ee-qwen-2-5-7b-p as the model parameter.Is the Qwen 2.5 7B API private?
Qwen 2.5 7B is end-to-end encrypted: prompts are encrypted on your device and decrypted only inside an attested hardware enclave, so neither Venice nor the GPU provider can read them.What is the context window of Qwen 2.5 7B?
32K tokens of context, with up to 4K output tokens per response.What does Qwen 2.5 7B support?
Qwen 2.5 7B supports web search.Which endpoint does the Qwen 2.5 7B API use?
CallPOST /chat/completions.Related models
- Qwen 3.8 27B API: 3.20 output per 1M tokens
- Qwen 3.6 27B API: 3.25 output per 1M tokens
- Qwen 3.5 9B API: 0.15 output per 1M tokens
- Qwen 3.5 397B API: 4.50 output per 1M tokens
- Mercury 2.5 API: 0.19 output per 1M tokens
- NVIDIA Nemotron 3 Nano 30B API: 0.30 output per 1M tokens
- Mistral Small 3.2 24B Instruct API: 0.25 output per 1M tokens
- OpenAI GPT OSS 120B API: 0.30 output per 1M tokens