Skip to main content

Gemini 3.7 Flash API

Gemini 3.7 Flash is a large language model by Google, available on the Venice API as gemini-3-7-flash. Requests are anonymized, so the provider never sees your identity.Gemini 3.7 Flash is Google’s most capable Flash model, built for complex coding, agentic workflows, and reliable multi-step execution, with 1M context and tunable thinking.

Gemini 3.7 Flash API pricing

Gemini 3.7 Flash specifications

How to use the Gemini 3.7 Flash API

Send requests to POST https://api.venice.ai/api/v1/chat/completions with "model": "gemini-3-7-flash" and your API key.

Gemini 3.7 Flash API FAQ

How much does the Gemini 3.7 Flash API cost?

0.94per1Minputtokensand0.94 per 1M input tokens and 4.69 per 1M output tokens, with cached input at $0.094 per 1M. Prices are in USD and can be paid in DIEM at parity.

What is the Gemini 3.7 Flash model ID?

Use gemini-3-7-flash as the model parameter.

Is the Gemini 3.7 Flash API private?

Gemini 3.7 Flash is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.

What is the context window of Gemini 3.7 Flash?

1M tokens of context, with up to 64K output tokens per response.

What does Gemini 3.7 Flash support?

Gemini 3.7 Flash supports function calling, structured outputs, reasoning, image input, web search and prompt caching.

Which endpoint does the Gemini 3.7 Flash API use?

Call POST /chat/completions. /responses (Alpha) is also supported.

Related models