Skip to main content

Grok 4.20 Multi-Agent API

Grok 4.20 Multi-Agent is a large language model by xAI, available on the Venice API as grok-4-20-multi-agent. It runs privately, with zero data retention.Grok 4.20 Multi-Agent is a variant of xAI Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information across complex tasks.

Grok 4.20 Multi-Agent API pricing

Grok 4.20 Multi-Agent specifications

How to use the Grok 4.20 Multi-Agent API

Send requests to POST https://api.venice.ai/api/v1/chat/completions with "model": "grok-4-20-multi-agent" and your API key.

Grok 4.20 Multi-Agent API FAQ

How much does the Grok 4.20 Multi-Agent API cost?

1.42per1Minputtokensand1.42 per 1M input tokens and 2.83 per 1M output tokens, with cached input at $0.23 per 1M. Prices are in USD and can be paid in DIEM at parity.

What is the Grok 4.20 Multi-Agent model ID?

Use grok-4-20-multi-agent as the model parameter.

Is the Grok 4.20 Multi-Agent API private?

Grok 4.20 Multi-Agent is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.

What is the context window of Grok 4.20 Multi-Agent?

2M tokens of context, with up to 125K output tokens per response.

What does Grok 4.20 Multi-Agent support?

Grok 4.20 Multi-Agent supports structured outputs, reasoning, image input, web search and prompt caching.

Which endpoint does the Grok 4.20 Multi-Agent API use?

Call POST /chat/completions. /responses (Alpha) is also supported.

Related models

  • GLM 5 Turbo API: 1.20input/1.20 input / 4.00 output per 1M tokens
  • Qwen 3.5 397B API: 0.75input/0.75 input / 4.50 output per 1M tokens
  • GLM 5 API: 1.00input/1.00 input / 3.20 output per 1M tokens
  • GPT-5.4 Mini API: 0.94input/0.94 input / 5.63 output per 1M tokens