Plain-text specification
Plain-text specification
GLM 5.2 API
GLM 5.2 is a large language model by Z.ai, available on the Venice API aszai-org-glm-5-2. Private and end-to-end encrypted variants are available.GLM-5.2 is the next-generation large language model developed by Zhiyuan AI, featuring significantly enhanced reasoning capabilities, improved instruction following, and support for multiple languages. Supports large context windows for processing extensive text and detailed analysis with fast inference speed.GLM 5.2 API pricing
GLM 5.2 specifications
How to use the GLM 5.2 API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "zai-org-glm-5-2" and your API key.GLM 5.2 API FAQ
How much does the GLM 5.2 API cost?
4.40 per 1M output tokens, with cached input at 1.75 input and $5.75 output. Prices are in USD and can be paid in DIEM at parity.What is the GLM 5.2 model ID?
Usezai-org-glm-5-2 as the model parameter. Other variants: e2ee-glm-5-2-p (E2EE).Is the GLM 5.2 API private?
The Standard variant is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training. The E2EE variant is end-to-end encrypted: prompts are encrypted on your device and decrypted only inside an attested hardware enclave, so neither Venice nor the GPU provider can read them.What is the context window of GLM 5.2?
1M tokens of context, with up to 128K output tokens per response.What does GLM 5.2 support?
GLM 5.2 supports function calling, structured outputs, reasoning, web search and prompt caching. Reasoning effort is adjustable withreasoning_effort: none, high and max (default max).Which endpoint does the GLM 5.2 API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- GLM 5.3 API: 5.50 output per 1M tokens
- GLM 5.1 API: 4.84 output per 1M tokens
- GLM 5 API: 3.20 output per 1M tokens
- GLM 4.7 API: 2.65 output per 1M tokens
- GLM 4.6 API: 1.75 output per 1M tokens
- Inkling API: 5.06 output per 1M tokens
- DeepSeek V4 Pro API: 3.30 output per 1M tokens
- GPT-5.4 Mini API: 5.63 output per 1M tokens
- Gemini 3.6 Flash API: 4.69 output per 1M tokens