Plain-text specification
Plain-text specification
Kimi K2.7 Code API
Kimi K2.7 Code is a large language model by Moonshot AI, available on the Venice API askimi-k2-7-code. It runs privately, with zero data retention.Kimi K2.7 Code is Moonshot AI’s coding-focused agentic model built on Kimi K2.6, with 1T total parameters and 32B active parameters. It always operates in thinking mode, supports text and image input, and targets long-horizon software engineering, agentic task decomposition, and multi-turn coding workflows with 256K context.Kimi K2.7 Code API pricing
Kimi K2.7 Code specifications
How to use the Kimi K2.7 Code API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "kimi-k2-7-code" and your API key.Kimi K2.7 Code API FAQ
How much does the Kimi K2.7 Code API cost?
3.50 per 1M output tokens, with cached input at $0.16 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the Kimi K2.7 Code model ID?
Usekimi-k2-7-code as the model parameter.Is the Kimi K2.7 Code API private?
Kimi K2.7 Code is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.What is the context window of Kimi K2.7 Code?
250K tokens of context, with up to 64K output tokens per response.What does Kimi K2.7 Code support?
Kimi K2.7 Code supports function calling, structured outputs, reasoning, image input, web search and prompt caching. Reasoning effort is adjustable withreasoning_effort: low, medium and high (default high).Which endpoint does the Kimi K2.7 Code API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- Qwen 3.6 Plus Uncensored API: 3.75 output per 1M tokens
- NVIDIA Nemotron 3 Ultra API: 3.13 output per 1M tokens
- Seed 2.1 Turbo API: 3.13 output per 1M tokens
- Grok Build 0.1 API: 2.00 output per 1M tokens