Skip to main content

GPT-4o Mini API

GPT-4o Mini is a large language model by OpenAI, available on the Venice API as openai-gpt-4o-mini-2024-07-18. Requests are anonymized, so the provider never sees your identity.OpenAI’s cost-efficient small model that delivers GPT-4 level intelligence at a fraction of the cost. Ideal for high-volume applications requiring strong reasoning. Version: 2024-07-18.

GPT-4o Mini API pricing

GPT-4o Mini specifications

How to use the GPT-4o Mini API

Send requests to POST https://api.venice.ai/api/v1/chat/completions with "model": "openai-gpt-4o-mini-2024-07-18" and your API key.

GPT-4o Mini API FAQ

How much does the GPT-4o Mini API cost?

0.19per1Minputtokensand0.19 per 1M input tokens and 0.75 per 1M output tokens, with cached input at $0.094 per 1M. Prices are in USD and can be paid in DIEM at parity.

What is the GPT-4o Mini model ID?

Use openai-gpt-4o-mini-2024-07-18 as the model parameter.

Is the GPT-4o Mini API private?

GPT-4o Mini is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.

What is the context window of GPT-4o Mini?

125K tokens of context, with up to 16K output tokens per response.

What does GPT-4o Mini support?

GPT-4o Mini supports function calling, structured outputs, image input, web search and prompt caching.

Which endpoint does the GPT-4o Mini API use?

Call POST /chat/completions. /responses (Alpha) is also supported.

Related models