Gemini 3 Flash Preview

Gemini 3 Flash preview, 常规快速档

Model ID: gemini-3-flash-preview · Type: chat · Provider: Google

Endpoints: /v1/chat/completions · /v1/messages

Pricing

Input (per 1M tokens)$0.375 USD
Output (per 1M tokens)$2.25 USD
Cache read (per 1M tokens)$0.0375 USD

Capabilities

from openai import OpenAI

client = OpenAI(api_key="sk-...", base_url="https://api.router.ai/v1")
resp = client.chat.completions.create(
    model="gemini-3-flash-preview",
    messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)

FAQ

How much does Gemini 3 Flash Preview cost on AI中转站?

Gemini 3 Flash Preview (`gemini-3-flash-preview`) is billed per usage at $0.375/1M in · $2.25/1M out, in USD. Current pricing is always listed at https://ai-zzz.com/models/gemini-3-flash-preview.

How do I call Gemini 3 Flash Preview through AI中转站?

Send a request to https://api.router.ai/v1/v1/chat/completions with the header `Authorization: Bearer <your API key>` and `"model": "gemini-3-flash-preview"`. The API is OpenAI-compatible, so any OpenAI SDK works by changing base_url to https://api.router.ai/v1 — no other code change.

Which endpoints does Gemini 3 Flash Preview support?

Gemini 3 Flash Preview can be called on: /v1/chat/completions; /v1/messages.

What is Gemini 3 Flash Preview's context window?

Gemini 3 Flash Preview accepts up to 1,048,576 input tokens and can return up to 65,535 output tokens. Requests exceeding the input limit are rejected before reaching the model.

What can Gemini 3 Flash Preview do?

Gemini 3 Flash Preview supports: vision, function_calling, prompt_caching, audio_input, thinking, long_context, cache.

Who makes Gemini 3 Flash Preview?

Gemini 3 Flash Preview is a chat model from Google, available through the AI中转站 gateway with the same API key as every other model.

Call it through the AI中转站 OpenAI-compatible endpoint (API base: https://api.router.ai/v1). AI agents can discover and call every model on this gateway through MCP (https://mcp.router.ai/mcp) with no manual integration.

API reference · All models