Skip to main content

DeepSeek V4 Pro API

DeepSeek V4 Pro is a large language model by DeepSeek, available on the Venice API as deepseek-v4-pro. It runs privately, with zero data retention.DeepSeek V4 Pro is a 1.6T-parameter Mixture-of-Experts model with 49B active parameters and a 1M-token context window. Built for advanced reasoning, coding, and long-horizon agentic workflows with a hybrid attention system for efficient long-context processing.

DeepSeek V4 Pro API pricing

DeepSeek V4 Pro specifications

How to use the DeepSeek V4 Pro API

Send requests to POST https://api.venice.ai/api/v1/chat/completions with "model": "deepseek-v4-pro" and your API key.

DeepSeek V4 Pro API FAQ

How much does the DeepSeek V4 Pro API cost?

1.65per1Minputtokensand1.65 per 1M input tokens and 3.30 per 1M output tokens, with cached input at $0.33 per 1M. Prices are in USD and can be paid in DIEM at parity.

What is the DeepSeek V4 Pro model ID?

Use deepseek-v4-pro as the model parameter.

Is the DeepSeek V4 Pro API private?

DeepSeek V4 Pro is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.

What is the context window of DeepSeek V4 Pro?

1M tokens of context, with up to 32K output tokens per response.

What does DeepSeek V4 Pro support?

DeepSeek V4 Pro supports function calling, structured outputs, reasoning, web search and prompt caching. Reasoning effort is adjustable with reasoning_effort: none, low, medium and high (default low).

Which endpoint does the DeepSeek V4 Pro API use?

Call POST /chat/completions. /responses (Alpha) is also supported.

Related models