Plain-text specification
Plain-text specification
Gemini 3 Flash Preview API
Gemini 3 Flash Preview is a large language model by Google, available on the Venice API asgemini-3-flash-preview. Requests are anonymized, so the provider never sees your identity.Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi-turn chat, and coding assistance. It delivers near Pro level reasoning with substantially lower latency.Gemini 3 Flash Preview API pricing
Gemini 3 Flash Preview specifications
How to use the Gemini 3 Flash Preview API
Send requests toPOST https://api.venice.ai/api/v1/chat/completions with "model": "gemini-3-flash-preview" and your API key.Gemini 3 Flash Preview API FAQ
How much does the Gemini 3 Flash Preview API cost?
3.75 per 1M output tokens, with cached input at $0.07 per 1M. Prices are in USD and can be paid in DIEM at parity.What is the Gemini 3 Flash Preview model ID?
Usegemini-3-flash-preview as the model parameter.Is the Gemini 3 Flash Preview API private?
Gemini 3 Flash Preview is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.What is the context window of Gemini 3 Flash Preview?
250K tokens of context, with up to 64K output tokens per response.What does Gemini 3 Flash Preview support?
Gemini 3 Flash Preview supports function calling, structured outputs, reasoning, image input, web search and prompt caching. Reasoning effort is adjustable withreasoning_effort: none, low, medium and high (default high).Which endpoint does the Gemini 3 Flash Preview API use?
CallPOST /chat/completions. /responses (Alpha) is also supported.Related models
- GLM 5 API: 3.20 output per 1M tokens
- Kimi K2.5 API: 3.50 output per 1M tokens
- Kimi K2.6 API: 3.50 output per 1M tokens
- Qwen 3.6 Plus Uncensored API: 3.75 output per 1M tokens