25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Google / Gemini Flash / Gemini 3 Flash Preview

Gemini 3 Flash Preview

New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs

Line

Gemini Flash

Weights

API only

Released

2025-12-17

Context

1.05M

Input

$0.07 / 1M

Output

$0.43 / 1M

Max output

66K

Coverage

not covered yet

Overview

Gemini 3 Flash Preview is a new model in Google's gemini-flash line. It brings frontier-style multimodal reasoning to cheaper runs, according to the maker. It is a preview release dated 2025-12-17, succeeding Gemini 2.5 Flash. It has a context window of 1,048,576 tokens and a maximum output of 65,536 tokens. Input pricing starts at $0.07 per million tokens, output at $0.43 per million, and cached input at $0.025 per million. The weights are not open. It accepts audio, image, pdf, text, and video, and returns text. It supports reasoning mode and tool calling.

Specs

Released2025-12-17
LineGoogle · Gemini Flash
WeightsAPI only
Context1.05M tokens
Max output66K tokens
Inputaudio, image, pdf, text, video
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoff2025-01
API idgemini-3-flash-preview

Pricing

Input$0.07 / 1M tokens
Cached input$0.025 / 1M tokens
Output$0.43 / 1M tokens

Against Gemini 2.5 Flash: input $0.09 to $0.07, output $0.71 to $0.43, cached input $0.021 to $0.025.

Range across the hosts that serve it: up to $0.70 per 1M input tokens.

Strengths

  • 1M token context window for long inputs
  • 64K token max output for detailed answers
  • Accepts audio, image, pdf, text, video
  • Reasoning mode and tool calling enabled
  • Input from $0.07 per million tokens

Best for

  • Reach for it for multimodal reasoning on large inputs
  • Reach for it for processing long documents and videos
  • Reach for it for cost-sensitive high-volume inference
  • Reach for it for tool-augmented agent workflows

How to access

ProviderModel id
302.AIgemini-3-flash-preview
AIHubMixgemini-3-flash-preview
Abacusgemini-3-flash-preview
AnyAPIgemini-3-flash-preview
CrossModelgemini-3-flash-preview
DevPass (LLM Gateway)gemini-3-flash-preview
Eden AIgemini-3-flash-preview
FrogBotgemini-3-flash-preview
Googlegemini-3-flash-preview
Jiekou.AIgemini-3-flash-preview
Kilo Gatewaygemini-3-flash-preview
LLM Gatewaygemini-3-flash-preview

29 hosts serve this model at the price above · the maker's documentation

Gemini Flash: every version

VersionReleasedContextInput
Gemini 3.8 FlashCurrent2026-09-021.05M$0.75
Gemini 3.7 Flash2026-08-131.05M$0.75
Gemini 3.6 Flash2026-07-211.05M$0.375
Gemini 3.5 Flash2026-05-191.05M$0.1857
Gemini 3 Flash Preview2025-12-171.05M$0.07
Gemini 2.5 Flash2025-06-171.05M$0.09

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window?
1,048,576 tokens, same as the previous Gemini 2.5 Flash.
Does it support image input?
Yes, it accepts audio, image, pdf, text, and video, and returns text.
What is the price?
From $0.07 per million input tokens and $0.43 per million output tokens. Cached input is $0.025 per million.
ProprietaryReasoning1.05M context