Google's mid-tier Gemini 3 Flash preview with multimodal input, a 1M-token context, and $0.50/$3.00 per million pricing.

Key specifications

Capability
88.2
Context window
1.0M
Max output
65.5K
Input $/1M
$0.50
Output $/1M
$3.00
License
Proprietary

Inputs and outputs

Recorded modality support shows what this model can accept and produce. An undocumented modality is shown as unknown rather than assumed to be unsupported.

Accepts

  • Text input
  • Image input
  • Video input
  • Audio input
  • PDF input

Produces

  • Text output

Capabilities

ReasoningFunction callingVisionAudioJSON mode

What is Gemini 3 Flash Preview?

gemini 3 flash preview is Google's mid-tier model from the 3-generation Gemini line, sitting between the ultra-cheap Flash Lite tier and the expensive Pro tier on price, while landing right in the middle on capability too. It brought frontier-style multimodal reasoning down to a cheaper price point when it shipped, and it's the model to reach for when a task needs more than simple extraction but doesn't justify Pro's cost.

Pricing and context in practice

Input tokens cost $0.50 per million, output tokens $3.00 per million, and cached input runs $0.05 per million — a straightforward 10x discount on repeated content. The context window is 1,048,576 tokens with up to 65,536 tokens of output, so it can hold long documents or extended conversation history without needing to chunk input across multiple calls. For a workload that sends a large reference document on every request, caching that document once and reusing it across calls cuts the effective input cost by 90%.

Where it earns its place

  • Multimodal reasoning tasks — text, image, video, audio, and PDF input — that need more judgment than a pure classification model
  • Agent steps that combine function calling with a moderate reasoning load
  • Workloads priced too high for Pro but too demanding for the lightest Flash Lite tier
  • Cases where a large context window matters more than rock-bottom per-token cost

How it sits against the rest of Google's lineup

The capability score here is 88.2, a step below both Gemini 3.1 Flash Lite Preview (88.6) and Gemini 3.1 Pro Preview (88.5) — a useful reminder that newer and cheaper doesn't always mean weaker, and older doesn't always mean stronger. Priced at $0.50/$3.00, it sits between Flash Lite's $0.25/$1.50 and Pro's $2.00/$12.00, but on this capability metric alone it doesn't clearly justify the premium over Flash Lite. Worth an A/B test on your own task before assuming the middle price tier is the middle-quality tier.

The honest limitation

This is a preview release from the 3 generation, not the newer 3.1 line, so newer Gemini releases may have already superseded it on cost or capability for the same task. Output is text only, and there's no benchmark suite attached to this listing — the capability score is the only comparative number available, and it should be treated as a rough signal, not a ranking.

How to evaluate Gemini 3 Flash Preview

Use the recorded facts as a shortlist, then validate the model against representative inputs and production constraints.

Workload fit

Gemini 3 Flash Preview is categorized for Audio-text-to-text, Automatic speech recognition, Image-text-to-text, Image-to-text, Text generation, Text-to-text AI models, Video-text-to-text. Its current record accepts text, image, video, audio, pdf and produces text. Confirm file formats, preprocessing, and provider-specific request schemas before implementation.

Capacity and cost

The directory records 1.0M context and 65.5K maximum output. Listed token prices are $0.50 input and $3.00 output per 1M tokens. Treat missing values as unknown and recheck current commercial terms.

Operational behavior

Recorded output speed is — and time to first token is —. Hosting route, region, prompt length, concurrency, and provider load can materially change both measurements.

Evidence boundary

This page separates sourced model facts from incomplete fields. Benchmark evidence is displayed only when its source is linked and verified. Before choosing Gemini 3 Flash Preview, test task quality, tool reliability, safety behavior, data controls, rate limits, and total cost on the exact route you intend to use.

Strengths

  • + Registry-backed capabilities
  • + Provider: Google

Limitations

  • Draft record — editorial review required
  • Verify pricing and independent evaluations before approval

Frequently asked questions

Input tokens are $0.50 per million, output tokens are $3.00 per million, and cached input is $0.05 per million tokens.

Related approved models in Automatic speech recognition. Compare specifications and verify fit for your workload.

Google

Gemini 3.5 Flash Lite

Google's low-cost Gemini 3 model with a 1,048,576-token context window, vision support, and $0.30/$2.50 per million token pricing for volume workloads.

1.0M contextCompare

Google

Gemini 3.6 Flash

Google's mid-tier Gemini 3 model: 1,048,576-token context, multimodal input, and $1.50/$7.50 per million token pricing for production agent workloads.

1.0M contextCompare

Google

Gemini 3.1 Flash Lite

Google's cheapest Gemini 3 model, with a 1,048,576-token context window and $0.25/$1.50 per million token pricing for high-volume, cost-sensitive work.

1.0M contextCompare

Google

Gemini 3.5 Flash

Google's full Gemini 3 Flash model, with a 1,048,576-token context window and $1.50/$9.00 per million token pricing for production-grade agent work.

1.0M contextCompare

Google

Gemini Flash Latest

Google's fast, rolling-updated Gemini model with a 1M-token context window, reasoning, and tool calling at $1.50/$9.00 per million tokens.

1.0M contextCompare

Google

Gemini Flash-Lite Latest

Google's low-cost, high-volume Gemini model at $0.25/$1.50 per million tokens, with a 1M-token context window and reasoning support.

1.0M contextCompare