Gemini 3.6 Flash: Google's New Free Default Model

Gemini 3.6 Flash: Google's New Free Default Model

Google just released Gemini 3.6 Flash: faster, cheaper, and now the default in the app and AI Studio. Here's what the free tier actually gives you.

Too much jargon?→ Look it up in the glossary

Odds are you already talked to a new AI model this week without noticing. If you asked gemini.google.com a question in the last few days, the old model wasn't the one answering anymore – Gemini 3.6 Flash was. No update dialog, no asterisk, just quietly swapped in. Welcome to everyday AI in 2026.

What's new in Gemini 3.6 Flash?

Flash has always been Google's fast, cheap model line – built for everyday tasks, not for squeezing out the last percentage points of reasoning power. The new version stays true to that, but gets noticeably more efficient: for comparable tasks, Google says it needs about 17% fewer output tokens than its predecessor, 3.5 Flash. Fewer tokens means faster answers and a cheaper bill.

The context window stays at a generous 1 million tokens – roughly the length of a thick novel the model can keep track of in one go. Its knowledge cutoff moved up to March 2026, and on tasks like coding, long-context retrieval, and computer-use control, current benchmark numbers show a clear step up.

Google also rolled out two siblings alongside it: Gemini 3.5 Flash-Lite (even leaner, even cheaper) and a variant tuned for security tasks. And yes, a glimpse of Gemini 4 has already been teased – the cycle keeps turning before you've even gotten used to 3.6.

What's free — and what actually costs money?

Two paths are free right now:

  • Gemini app: Just keep chatting as usual. Gemini 3.6 Flash runs in the background as the new default for the free tier, no paid plan required.
  • Google AI Studio: If you want to experiment yourself or build small tools, grab a free API key there. No subscription, no credit card. In exchange, Google caps the free tier at a modest number of requests per minute and per day – plenty for trying things out and small side projects, not for running something in production.

Once things get serious, the API bills you by tokens: currently around $1.50 per million input tokens and $7.50 per million output tokens. Compared to the big flagship models, that's a fraction of the cost – Flash is deliberately economy class, not the front-row seat.

The catch with free access, as usual, is that Google reserves the right to use inputs from free-tier usage to improve its products. For everyday chats about recipes or travel plans, that's easy to live with. For sensitive business or customer data, it doesn't belong in the free tier – that's what the paid plan with stricter privacy terms is for.

Is this for me — or just for developers?

The Gemini app affects anyone with a Google account who enjoys an AI chat: there's nothing to do, the upgrade already happened. Using AI Studio, on the other hand, already leans toward "I'm building something myself." It's not rocket science – a Google account is enough, and the API key takes a few clicks – but it's aimed at people who aren't scared off by a text field surrounded by code.

What you can try right now

Open gemini.google.com, ask something that used to take a moment – a coding problem, a long PDF summary, a multi-step task – and see if it feels faster. For tinkerers: grab the free key in AI Studio and talk to the model gemini-3.6-flash directly, no app wrapper needed.

The real story here isn't really the single model, it's the pace behind it: every few weeks, whatever you're using for free gets swapped out under the hood, usually for the better. No need to get jumpy about it. But it's worth checking every once in a while what "the default model" even is, before you start taking it for granted.