Plain-text specification
Plain-text specification
Mistral Small 3.2 24B Instruct API
Mistral Small 3.2 24B Instruct is a large language model by Mistral AI, available on the Venice API asmistral-small-3-2-24b-instruct. It runs privately, with zero data retention.Mistral Small 3.2 is a 24B parameter model optimized for efficiency and performance. Ideal for general-purpose tasks with balanced speed and capability.Mistral Small 3.2 24B Instruct API pricing
Mistral Small 3.2 24B Instruct specifications
How to use the Mistral Small 3.2 24B Instruct API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "mistral-small-3-2-24b-instruct" and your API key.Mistral Small 3.2 24B Instruct API FAQ
How much does the Mistral Small 3.2 24B Instruct API cost?
0.25 per 1M output tokens. Prices are in USD and can be paid in DIEM at parity.What is the Mistral Small 3.2 24B Instruct model ID?
Usemistral-small-3-2-24b-instruct as the model parameter.Is the Mistral Small 3.2 24B Instruct API private?
Mistral Small 3.2 24B Instruct is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.What is the context window of Mistral Small 3.2 24B Instruct?
250K tokens of context, with up to 16K output tokens per response.What does Mistral Small 3.2 24B Instruct support?
Mistral Small 3.2 24B Instruct supports function calling, structured outputs, image input and web search.Which endpoint does the Mistral Small 3.2 24B Instruct API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- NVIDIA Nemotron 3 Nano 30B API: 0.30 output per 1M tokens
- GLM 4.7 Flash API: 0.40 output per 1M tokens
- OpenAI GPT OSS 120B API: 0.30 output per 1M tokens
- Google Gemma 3 27B Instruct API: 0.20 output per 1M tokens