Skip to main content

Gemini 3 Flash Preview API

Gemini 3 Flash Preview is a large language model by Google, available on the Venice API as gemini-3-flash-preview. Requests are anonymized, so the provider never sees your identity.Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi-turn chat, and coding assistance. It delivers near Pro level reasoning with substantially lower latency.

Gemini 3 Flash Preview API pricing

Gemini 3 Flash Preview specifications

How to use the Gemini 3 Flash Preview API

Send requests to POST https://api.venice.ai/api/v1/chat/completions with "model": "gemini-3-flash-preview" and your API key.

Gemini 3 Flash Preview API FAQ

How much does the Gemini 3 Flash Preview API cost?

0.70per1Minputtokensand0.70 per 1M input tokens and 3.75 per 1M output tokens, with cached input at $0.07 per 1M. Prices are in USD and can be paid in DIEM at parity.

What is the Gemini 3 Flash Preview model ID?

Use gemini-3-flash-preview as the model parameter.

Is the Gemini 3 Flash Preview API private?

Gemini 3 Flash Preview is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.

What is the context window of Gemini 3 Flash Preview?

250K tokens of context, with up to 64K output tokens per response.

What does Gemini 3 Flash Preview support?

Gemini 3 Flash Preview supports function calling, structured outputs, reasoning, image input, web search and prompt caching. Reasoning effort is adjustable with reasoning_effort: none, low, medium and high (default high).

Which endpoint does the Gemini 3 Flash Preview API use?

Call POST /chat/completions. /responses (Alpha) is also supported.

Related models