← Back to API models

Gemini 3.6 Flash

by Google DeepMind · workhorse Flash · released July 21, 2026

Google cut 3.6 Flash's price to $0.75/$3.75 on Aug 13, 2026, matching new sibling Gemini 3.7 Flash's intro rate through the end of the year.

Input (intro) $0.75 / 1M
Output (intro) $3.75 / 1M
Context 1M tokens
Reverts $1.50/$7.50 Jan 1, 2027
Google AI Studio ↗ Updated August 17, 2026
§ API pricing

Per-token rates, now on an introductory discount.

Input (intro)
$0.75/1M tokens
Prompt tokens
  • Cut from $1.50 on Aug 13, 2026
  • Vision, audio, video billed as input
  • Cached input just $0.075 / 1M
Output (intro)
$3.75/1M tokens
Completion tokens
  • Cut from $7.50 on Aug 13, 2026
  • Matches new sibling Gemini 3.7 Flash
  • Includes thinking tokens
Context
1Mtokens
Window · 64K out
  • Full 1M input, up to 64K output
  • Native multimodal across the window
  • Thinking controls + Computer Use
Reverts Jan 1, 2027
$1.50/$7.50 per 1M
Standard rate
  • Intro pricing holds through Dec 31, 2026
  • Same rate 3.6 Flash launched at in July
  • No action needed — automatic on Google's end

Gemini 3.6 Flash now bills $0.75 / 1M input and $3.75 / 1M output through Dec 31, 2026 — half its original launch price, cut on Aug 13, 2026 to match the newly launched Gemini 3.7 Flash. Cached input is $0.075 / 1M. The rate reverts to $1.50/$7.50 on Jan 1, 2027.

What's new: a price cut, not a model change

Gemini 3.6 Flash itself hasn't changed since its July 21, 2026 release — replacing Gemini 3.5 Flash with 17% fewer output tokens on the Artificial Analysis Index and built-in Computer Use. What changed on August 13, 2026 is the price: Google launched a newer sibling, Gemini 3.7 Flash, at an introductory $0.75/$3.75 per 1M tokens, and cut 3.6 Flash to the same rate rather than leave its own older model looking overpriced next to the new one. Both models now cost the same through the end of 2026.

The intro window is temporary: unless Google extends it, 3.6 Flash's price reverts to its original $1.50/$7.50 on January 1, 2027. Budget accordingly if you're building around the discounted rate for the long term.

Capabilities

Flash inherits the native multimodality of the Gemini family: text, image, audio, and video all go in, across the full 1M context, with up to 64K output tokens and thinking controls. The 3.6 generation leans hardest into coding and computer use, making it a strong fit for agent loops and IDE assistants. Gemini 3.7 Flash is now Google's more explicitly coding-and-agent-focused pick at the same price, so the choice between the two comes down to which benchmark profile fits your workload.

Typical use cases

  • Coding assistants and IDE integrations
  • Agentic UI automation via built-in Computer Use
  • Production chat where quality and cost both matter
  • Long-document and video/audio analysis over 1M context
  • High-volume workloads while the intro price lasts

Sibling and rival comparison

ModelInput / 1MOutput / 1MContext
Gemini 3.6 Flash (intro)$0.75$3.751M
Gemini 3.7 Flash (intro)$0.75$3.751M
Gemini 3.5 Flash-Lite$0.30$2.501M
Gemini 3.1 Pro$2$121M
GPT-5.6 Luna$0.20$1.201M

3.6 Flash and 3.7 Flash are priced identically through 2026 — pick 3.7 Flash if you want Google's newest coding/agent tuning, or stick with 3.6 Flash if you already have it validated in production. Flash-Lite remains the genuine cheap-volume tier, and GPT-5.6 Luna still undercuts both Flash models on raw price but lacks native video.

← See all Google / Gemini plans