Plain-text specification
Plain-text specification
Gemma 4 Uncensored API
Gemma 4 Uncensored is a large language model by Google, available on the Venice API asgemma-4-uncensored. It runs privately, with zero data retention.Gemma 4 Uncensored is an uncensored variant of Google Gemma 4 26B, a Mixture-of-Experts model with 26B total parameters and only 4B active per token. Fine-tuned for uncensored chat without content filtering, it supports 256K context, coding, and general-purpose conversation.Gemma 4 Uncensored API pricing
Gemma 4 Uncensored specifications
How to use the Gemma 4 Uncensored API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "gemma-4-uncensored" and your API key.Gemma 4 Uncensored API FAQ
How much does the Gemma 4 Uncensored API cost?
0.50 per 1M output tokens. Prices are in USD and can be paid in DIEM at parity.What is the Gemma 4 Uncensored model ID?
Usegemma-4-uncensored as the model parameter.Is the Gemma 4 Uncensored API private?
Gemma 4 Uncensored is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.What is the context window of Gemma 4 Uncensored?
250K tokens of context, with up to 8K output tokens per response.What does Gemma 4 Uncensored support?
Gemma 4 Uncensored supports function calling, structured outputs, image input and web search.Which endpoint does the Gemma 4 Uncensored API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- GLM 5.3 Flash API: 0.50 output per 1M tokens
- GPT-6 Luna API: 0.63 output per 1M tokens
- DeepSeek V4 Flash 0731 API: 0.35 output per 1M tokens
- Qwen 3.8 Flash API: 0.49 output per 1M tokens