Skip to content
GoogleChat

Gemini 3 Flash Preview

gemini-3-flash-preview

Gemini 3 Flash Preview is a fast, cost-efficient reasoning model designed for agent workflows, multi-turn chat, and coding assistance. It offers near-Pro level reasoning and tool-use performance with significantly lower latency than larger Gemini models, making it ideal for interactive development and long-running agent loops. It improves on Gemini 2.5 Flash with stronger reasoning, multimodal understanding, and reliability. The model supports a 1M-token context window and multimodal inputs (text, images, audio, video, PDFs) with text outputs. It provides configurable reasoning levels, structured output formats, tool use, and automatic context caching—optimized for users seeking strong agentic reasoning without the cost or latency of full frontier-scale models.

Context
1.0M tokens
Endpoint

Service Status

Status information temporarily unavailable

Apertis cannot confirm the current service state. This is not a report that the model is down.

Get API KeyCompare

Pricing

Input$0.25 / 1M
Output$1.50 / 1M
Cache Write (5m)$0.25 / 1M
Cache Write (1h)$0.25 / 1M
Cache Read$0.25 / 1M
Web Search$0 / 1M

Quick Start

Select an endpoint and copy a working example for this model.

Endpoint
python
from openai import OpenAI client = OpenAI(    api_key="YOUR_API_KEY",    base_url="https://api.apertis.ai/v1") response = client.chat.completions.create(    model="gemini-3-flash-preview",    messages=[        {"role": "user", "content": "Hello!"}    ],    max_tokens=1024,    temperature=0.7) print(response.choices[0].message.content) # Optional: Enable context compression to reduce token usage# response = client.chat.completions.create(#     model="gemini-3-flash-preview",#     messages=[{"role": "user", "content": "Hello!"}],#     extra_body={"compression": {"enabled": True, "model": "gpt-4.1-mini"}}# )

Supported Parameters

API docs
Common7 params
modelmessagesmax_tokenstemperaturetop_pstreamtools
Extended4 params
reasoning_effortstream_optionsthinkingextra_body

Cursor IDE Model IDs

Use these namespaced identifiers in Cursor IDE to avoid conflicts with built-in models.

gemini-3-flash-preview

Compare with Other Models

See how this model compares to others from the same provider.