Plain-text specification
Plain-text specification
MiniMax M2.7 API
MiniMax M2.7 is a large language model by MiniMax, available on the Venice API asminimax-m27. It runs privately, with zero data retention.MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity with advanced agentic capabilities through multi-agent collaboration.MiniMax M2.7 API pricing
MiniMax M2.7 specifications
How to use the MiniMax M2.7 API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "minimax-m27" and your API key.MiniMax M2.7 API FAQ
How much does the MiniMax M2.7 API cost?
1.50 per 1M output tokens, with cached input at $0.069 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the MiniMax M2.7 model ID?
Useminimax-m27 as the model parameter.Is the MiniMax M2.7 API private?
MiniMax M2.7 is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.What is the context window of MiniMax M2.7?
198K tokens of context, with up to 32K output tokens per response.What does MiniMax M2.7 support?
MiniMax M2.7 supports function calling, reasoning, web search and prompt caching. Reasoning effort is adjustable withreasoning_effort: none, low, medium and high (default low).Which endpoint does the MiniMax M2.7 API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- MiniMax M2.5 API: 0.95 output per 1M tokens
- Qwen 3 Coder 480B Turbo API: 1.50 output per 1M tokens
- Qwen3 VL 235B API: 1.90 output per 1M tokens
- Qwen 3.5 35B A3B API: 1.25 output per 1M tokens
- DeepSeek V4.1 Flash API: 1.50 output per 1M tokens