25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Google / Gemini Flash / Gemini 3.6 Flash

Gemini 3.6 Flash

Fast Gemini model balancing multimodal reasoning, tool use, and cost

Line

Gemini Flash

Weights

API only

Released

2026-07-21

Context

1.05M

Input

$0.375 / 1M

Output

$1.88 / 1M

Max output

66K

Coverage

1 signal

Overview

Gemini 3.6 Flash is a fast Gemini model that balances multimodal reasoning, tool use, and cost. It is the latest in the gemini-flash line, succeeding Gemini 3.5 Flash. It has a context window of 1,048,576 tokens and a maximum output of 65,536 tokens. It accepts audio, image, pdf, text, and video, and returns text. It supports reasoning mode and tool calling. The weights are not open. Pricing starts at $0.375 per million input tokens and $1.88 per million output tokens, with cached input at $0.0375 per million tokens.

Specs

Released2026-07-21
LineGoogle · Gemini Flash
WeightsAPI only
Context1.05M tokens
Max output66K tokens
Inputaudio, image, pdf, text, video
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoff2026-03
API idgemini-3.6-flash

Pricing

Input$0.375 / 1M tokens
Cached input$0.0375 / 1M tokens
Output$1.88 / 1M tokens

Against Gemini 3.5 Flash: input $0.1857 to $0.375, output $1.11 to $1.88, cached input $0.075 to $0.0375.

Range across the hosts that serve it: up to $1.50 per 1M input tokens.

Strengths

  • Fast multimodal reasoning and tool use
  • 1,048,576 token context window
  • Accepts audio, image, pdf, text, video
  • Supports reasoning mode and tool calling
  • Max output of 65,536 tokens

Best for

  • Reach for it for fast multimodal reasoning and tool use
  • Reach for it for processing inputs up to 1M tokens
  • Reach for it for cost-sensitive applications with cached input

How to access

ProviderModel id
302.AIgemini-3.6-flash
AIHubMixgemini-3.6-flash
Abacusgemini-3.6-flash
Cortecsgemini-3.6-flash
CrossModelgemini-3.6-flash
DevPass (LLM Gateway)gemini-3.6-flash
Eden AIgemini-3.6-flash
GitHub Copilotgemini-3.6-flash
Googlegemini-3.6-flash
Impossiblgemini-3.6-flash
Kilo Gatewaygemini-3.6-flash
LLM Gatewaygemini-3.6-flash

23 hosts serve this model at the price above · the maker's documentation

What the digest said

1 signal of the archive mention this model, each scored against the published bar. This list is rebuilt from the archive at every build, so it cannot go stale.

Gemini Flash: every version

VersionReleasedContextInput
Gemini 3.8 FlashCurrent2026-09-021.05M$0.75
Gemini 3.7 Flash2026-08-131.05M$0.75
Gemini 3.6 Flash2026-07-211.05M$0.375
Gemini 3.5 Flash2026-05-191.05M$0.1857
Gemini 3 Flash Preview2025-12-171.05M$0.07
Gemini 2.5 Flash2025-06-171.05M$0.09

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window?
The context window is 1,048,576 tokens, and the maximum output is 65,536 tokens.
What are the pricing details?
Input starts at $0.375 per million tokens, output at $1.88 per million, and cached input at $0.0375 per million.
What input types does it accept?
It accepts audio, image, pdf, text, and video, and returns text.
ProprietaryReasoning1.05M context