Gemini 3.6 Flash

GA

GA/stable ('Stable' on ai.google.dev/gemini-api/docs/models, described there as Google's latest Flash). Added 2026-07-31 after Google's deprecations table began naming it as the migration target for gemini-2.5-flash and the gemini-2.0-flash line. Prices are the paid-tier STANDARD rates from the official pricing page. CORRECTED 2026-08-14: the fields here carried Google's January-2027 column ($1.50/$7.50/$0.15) rather than the rate actually in force. Google prices this model on a two-date schedule: $0.75 in / $3.75 out per 1M with context caching $0.075/1M 'through December 31, 2026', doubling to $1.50 / $7.50 / $0.15 'starting January 1, 2027'. The per-1M-tokens-per-hour cache STORAGE fee this schema has no field for follows the same schedule — $0.50 through 2026-12-31, $1.00 from 2027-01-01 — and the note here previously gave the 2027 figure for that too. Other published tiers, none of which fit the per-MTok fields either, all on the same two dates: Batch and Flex $0.375/$1.875 (caching $0.0375) then $0.75/$3.75 (caching $0.075); Priority $1.35/$6.75 (caching $0.135) then $2.70/$13.50 (caching $0.27). Grounding with Google Search or Maps: 5,000 requests/month free shared across all Gemini 3.x models, then $14 per 1,000. Spec table gives 1,048,576 input / 65,536 output tokens and inputs text, image, video, audio, PDF. CORRECTED 2026-08-02: the note here previously asserted Google publishes no release date. That was true of the two surfaces it was read from (model page, pricing page) and false about the world — Google's DEPRECATIONS table carries it as data, 'gemini-3.6-flash | July 21, 2026 | No shutdown date announced', and the API changelog entry of 2026-07-21 says the same ('Gemini 3.6 Flash and Gemini 3.5 Flash-Lite generally available'). released is now 2026-07-21. The knowledge cutoff genuinely is unpublished on all three surfaces and stays null; the model page's 'Latest update July 2026' is a docs-edit date, not a launch. While the promotional window runs it is cheaper than Gemini 3.5 Flash on BOTH axes, not just output — 3.5 Flash is a flat $1.50 in / $9.00 out with no 2026/2027 split on its own block — so the two Flash rows are not a straight ladder, and the gap closes on 2027-01-01.

Gemini 3.6 Flash by Google costs $0.75 per 1M input tokens and $3.75 per 1M output tokens ($1.50/1M blended), with a 1.0M (1.048.576-token) context window. It is generally available (GA).

Last verified: 19 Aug 2026 · sourced from official provider documentation

Provider
Google
Status
GA
Input price
$0.75 / 1M tokens
Output price
$3.75 / 1M tokens
Cached input
$0.075 / 1M tokens
Blended price
$1.50 / 1M tokens
Context window
1.048.576 tokens (1.0M)
Max output
65.536 tokens
Modality
text, image, video, audio, pdf
Open weights
No — API only
Knowledge cutoff
Released
21 Jul 2026
API string
gemini-3.6-flash

Source: Google official documentation ↗

FAQ

Gemini 3.6 Flash — questions & answers

How much does Gemini 3.6 Flash cost?

Gemini 3.6 Flash is priced at $0.75 per 1M input tokens and $3.75 per 1M output tokens ($1.50/1M blended at a 3:1 input-to-output ratio), with cached input at $0.075 per 1M tokens.

What is the context window of Gemini 3.6 Flash?

Gemini 3.6 Flash has a 1.048.576-token context window (1.0M), with up to 65.536 output tokens per request.

Is Gemini 3.6 Flash deprecated?

No — Gemini 3.6 Flash is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.

Does Gemini 3.6 Flash support image or vision input?

Yes — Gemini 3.6 Flash accepts image input. Listed modalities: text, image, video, audio, pdf.

Can I self-host Gemini 3.6 Flash?

No — Gemini 3.6 Flash is proprietary. Its weights are not publicly released, so it is available only through Google's API (or hosted partners), not for self-hosting.

What is the API model string for Gemini 3.6 Flash?

The API model identifier for Gemini 3.6 Flash is "gemini-3.6-flash" when calling Google's API.