Plain-text specification
Plain-text specification
Grok 4.20 API
Grok 4.20 is a large language model by xAI, available on the Venice API asgrok-4-20. It runs privately, with zero data retention.Grok 4.20 is xAI’s latest multimodal reasoning model with strong tool use, structured output support, and a 2M-token context window.Grok 4.20 API pricing
Grok 4.20 specifications
How to use the Grok 4.20 API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "grok-4-20" and your API key.Grok 4.20 API FAQ
How much does the Grok 4.20 API cost?
2.83 per 1M output tokens, with cached input at $0.23 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the Grok 4.20 model ID?
Usegrok-4-20 as the model parameter.Is the Grok 4.20 API private?
Grok 4.20 is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.What is the context window of Grok 4.20?
2M tokens of context, with up to 125K output tokens per response.What does Grok 4.20 support?
Grok 4.20 supports function calling, structured outputs, reasoning, image input, web search and prompt caching.Which endpoint does the Grok 4.20 API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- Grok 4.7 API: 6.80 output per 1M tokens
- Grok 4.6 API: 6.80 output per 1M tokens
- Grok 4.5 API: 6.80 output per 1M tokens
- Grok 4.3 API: 2.83 output per 1M tokens
- GLM 5 Turbo API: 4.00 output per 1M tokens
- Qwen 3.5 397B API: 4.50 output per 1M tokens
- GLM 5 API: 3.20 output per 1M tokens
- GPT-5.4 Mini API: 5.63 output per 1M tokens