Plain-text specification
Plain-text specification
Gemini 3.6 Flash API
Gemini 3.6 Flash is a large language model by Google, available on the Venice API asgemini-3-6-flash. Requests are anonymized, so the provider never sees your identity.Gemini 3.6 Flash is a high speed, high value thinking model with 1M context, designed for agentic workflows, multi-turn chat, and coding assistance. It delivers near Pro level reasoning with substantially lower latency.Gemini 3.6 Flash API pricing
Gemini 3.6 Flash specifications
How to use the Gemini 3.6 Flash API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "gemini-3-6-flash" and your API key.Gemini 3.6 Flash API FAQ
How much does the Gemini 3.6 Flash API cost?
4.69 per 1M output tokens, with cached input at $0.094 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the Gemini 3.6 Flash model ID?
Usegemini-3-6-flash as the model parameter.Is the Gemini 3.6 Flash API private?
Gemini 3.6 Flash is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.What is the context window of Gemini 3.6 Flash?
1M tokens of context, with up to 64K output tokens per response.What does Gemini 3.6 Flash support?
Gemini 3.6 Flash supports function calling, structured outputs, reasoning, image input, web search and prompt caching.Which endpoint does the Gemini 3.6 Flash API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- Gemini 3.8 Flash API: 4.69 output per 1M tokens
- Gemini 3.7 Flash API: 4.69 output per 1M tokens
- Gemini 3.5 Flash API: 9.45 output per 1M tokens
- GLM 5.2 API: 4.40 output per 1M tokens
- Grok 4.3 API: 2.83 output per 1M tokens
- Inkling API: 5.06 output per 1M tokens
- GLM 5 Turbo API: 4.00 output per 1M tokens