Plain-text specification
Plain-text specification
DeepSeek V3.2 API
DeepSeek V3.2 is a large language model by DeepSeek, available on the Venice API asdeepseek-v3.2. It runs privately, with zero data retention.DeepSeek-V3.2 is an efficient large language model with DeepSeek Sparse Attention (DSA) for long contexts. It features strong reasoning and tool-use skills, achieving top results on the 2025 IMO and IOI.DeepSeek V3.2 API pricing
DeepSeek V3.2 specifications
How to use the DeepSeek V3.2 API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "deepseek-v3.2" and your API key.DeepSeek V3.2 API FAQ
How much does the DeepSeek V3.2 API cost?
0.48 per 1M output tokens, with cached input at $0.16 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the DeepSeek V3.2 model ID?
Usedeepseek-v3.2 as the model parameter.Is the DeepSeek V3.2 API private?
DeepSeek V3.2 is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.What is the context window of DeepSeek V3.2?
160K tokens of context, with up to 32K output tokens per response.What does DeepSeek V3.2 support?
DeepSeek V3.2 supports function calling, structured outputs, reasoning, web search and prompt caching. Reasoning effort is adjustable withreasoning_effort: none, low, medium and high (default low).Which endpoint does the DeepSeek V3.2 API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- Venice Uncensored 1.2 API: 0.90 output per 1M tokens
- GPT-4o Mini API: 0.75 output per 1M tokens
- Gemma 4 26B A4B Uncensored API: 0.88 output per 1M tokens
- Qwen3 VL 30B A3B API: 0.90 output per 1M tokens