Skip to main content

MiniMax M3 Preview API

MiniMax M3 Preview is a large language model by MiniMax, available on the Venice API as minimax-m3-preview. It runs privately, with zero data retention.MiniMax-M3 preview is a 1.4T-parameter frontier model from MiniMax for coding, agentic workflows, and complex reasoning, served at fp8 with a 512K context window.

MiniMax M3 Preview API pricing

MiniMax M3 Preview specifications

How to use the MiniMax M3 Preview API

Send requests to POST https://api.venice.ai/api/v1/chat/completions with "model": "minimax-m3-preview" and your API key.

MiniMax M3 Preview API FAQ

How much does the MiniMax M3 Preview API cost?

0.30per1Minputtokensand0.30 per 1M input tokens and 1.20 per 1M output tokens, with cached input at $0.06 per 1M. Prices are in USD and can be paid in DIEM at parity.

What is the MiniMax M3 Preview model ID?

Use minimax-m3-preview as the model parameter.

Is the MiniMax M3 Preview API private?

MiniMax M3 Preview is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.

What is the context window of MiniMax M3 Preview?

512K tokens of context, with up to 64K output tokens per response.

What does MiniMax M3 Preview support?

MiniMax M3 Preview supports function calling, reasoning, image input, web search and prompt caching.

Which endpoint does the MiniMax M3 Preview API use?

Call POST /chat/completions. /responses (Alpha) is also supported.

Related models