Plain-text specification
Plain-text specification
Grok 4.20 Multi-Agent API
Grok 4.20 Multi-Agent is a large language model by xAI, available on the Venice API asgrok-4-20-multi-agent. It runs privately, with zero data retention.Grok 4.20 Multi-Agent is a variant of xAI Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information across complex tasks.Grok 4.20 Multi-Agent API pricing
Grok 4.20 Multi-Agent specifications
How to use the Grok 4.20 Multi-Agent API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "grok-4-20-multi-agent" and your API key.Grok 4.20 Multi-Agent API FAQ
How much does the Grok 4.20 Multi-Agent API cost?
2.83 per 1M output tokens, with cached input at $0.23 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the Grok 4.20 Multi-Agent model ID?
Usegrok-4-20-multi-agent as the model parameter.Is the Grok 4.20 Multi-Agent API private?
Grok 4.20 Multi-Agent is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.What is the context window of Grok 4.20 Multi-Agent?
2M tokens of context, with up to 125K output tokens per response.What does Grok 4.20 Multi-Agent support?
Grok 4.20 Multi-Agent supports structured outputs, reasoning, image input, web search and prompt caching.Which endpoint does the Grok 4.20 Multi-Agent API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- GLM 5 Turbo API: 4.00 output per 1M tokens
- Qwen 3.5 397B API: 4.50 output per 1M tokens
- GLM 5 API: 3.20 output per 1M tokens
- GPT-5.4 Mini API: 5.63 output per 1M tokens