GEMINI LABJP
SUNSET — The image generation models shut down tomorrow, August 17: imagen-4.0-generate-001, ultra, fast, and the Gemini 3 Image family, and calls will fail with a hard errorGA — Gemini 3.7 Flash reached general availability on August 13, with substantial gains in software engineering, web development, and agentic work at an introductory price through December 31APPS — On August 12 Google widened the set of apps you can connect to Gemini, adding Granola, Otter.ai, and Wix alongside OpenTable, Ticketmaster, iHeartRadio, and PandoraSAMPLING — The temperature, top_p, and top_k sampling parameters are now deprecated, so migrating to a newer model means revisiting those assumptionsROBOTICS — Gemini Robotics ER 2 is in public preview, and the older gemini-robotics-er-1.6-preview shuts down on August 31NOTEBOOK — NotebookLM Enterprise has been renamed Gemini Notebook Enterprise, and the Gemini Enterprise mobile app is now generally availableSUNSET — The image generation models shut down tomorrow, August 17: imagen-4.0-generate-001, ultra, fast, and the Gemini 3 Image family, and calls will fail with a hard errorGA — Gemini 3.7 Flash reached general availability on August 13, with substantial gains in software engineering, web development, and agentic work at an introductory price through December 31APPS — On August 12 Google widened the set of apps you can connect to Gemini, adding Granola, Otter.ai, and Wix alongside OpenTable, Ticketmaster, iHeartRadio, and PandoraSAMPLING — The temperature, top_p, and top_k sampling parameters are now deprecated, so migrating to a newer model means revisiting those assumptionsROBOTICS — Gemini Robotics ER 2 is in public preview, and the older gemini-robotics-er-1.6-preview shuts down on August 31NOTEBOOK — NotebookLM Enterprise has been renamed Gemini Notebook Enterprise, and the Gemini Enterprise mobile app is now generally available
Articles/API / SDK
API / SDK/2026-08-16Intermediate

Moving to the Batch Tier Cancels Out the Gemini 3.7 Flash Price Increase Exactly

Introductory pricing for Gemini 3.7 Flash ends on December 31, 2026, and the rate doubles the next day. Here is which models are affected, which are not, and how to project your January bill from the usage you already have.

Gemini API211Pricing4Cost designModel selectionIndie development2

Premium Article

Gemini 3.7 Flash went generally available on August 13. The announcement mentioned introductory pricing through December 31, 2026, so I opened the pricing page to see how much of a discount that was.

The numbers were identical to Gemini 3.6 Flash.

$0.75 in, $3.75 out. Not a discount attached to the new model — the same deadline applies across the Flash line. And a little further down the page, the story shifted. This is not a discount that expires. It is a rate that doubles.

Some models double. Some do not. If you cross into January with that line blurred, the invoice will draw it for you.

The introductory rate is not exclusive to 3.7 Flash

Here are the published numbers. Checked on August 16, 2026 against the Gemini Developer API pricing page, in USD per 1M tokens, paid tier, Standard.

ModelInput (2026)Input (2027+)Output (2026)Output (2027+)
gemini-3.7-flash$0.75$1.50$3.75$7.50
gemini-3.6-flash$0.75$1.50$3.75$7.50
gemini-3.5-flash$1.50$1.50$9.00$9.00
gemini-3.5-flash-lite$0.30$0.30$2.50$2.50

3.7 Flash and 3.6 Flash match on both input and output. So the usual assumption — newer means pricier — does not hold here. There is no cost argument against moving from 3.6 to 3.7.

The row that surprised me was 3.5 Flash. An older generation, at twice the rate of 3.7 Flash, with no introductory pricing at all, so it stays at $1.50 / $9.00 into next year. Put that next to 3.7 Flash after the change ($1.50 / $7.50) and the input is level while the output is cheaper on 3.7. There is no financial reason to sit on 3.5 Flash, this year or next.

If you have been deferring the move to a newer model until things settle down, that deferral is not saving you anything.

Some models rise, some hold

Look again at the right half of that table. Only 3.7 Flash and 3.6 Flash move. 3.5 Flash and 3.5 Flash-Lite hold.

This is not an across-the-board increase. It changes the ratio between models.

Comparison20262027+
3.7 Flash ÷ 3.5 Flash-Lite (input)2.5×5.0×
3.7 Flash ÷ 3.5 Flash-Lite (output)1.5×3.0×

If you route between a light model and a heavy one, that shift lands directly on your design. A "when in doubt, escalate" policy that is affordable today may not be affordable in January. I will put numbers on the escalation rate later in this article, but the point to hold onto is that the increase acts on your architecture, not on a single model.

There is a second reading of the same fact. Because Flash-Lite holds, the incentive to push work down to it gets stronger next year. If quality allows, the next few months are a good window to widen the range of tasks Flash-Lite can handle.

Thank you for reading this far.

Continue Reading

What follows includes implementation code, benchmarks, and practical content we hope you'll find useful. This site runs without ads — server and development costs are supported entirely by members like you. If it's been helpful, we'd be truly grateful for your support.

WHAT YOU'LL LEARN
You will be able to project how much your Gemini spend rises on January 1, 2027 from your current usage, without waiting for the invoice
You will be able to tell which models are affected by the increase and which are not, so you are not hunting for a migration target in January
You will be able to separate the work that can wait from the work that cannot, and assign each to the Standard or Batch tier on your own criteria
Secure payment via Stripe · Cancel anytime

Unlock This Article

Get full access to the rest of this article. Buy once, read anytime. This site is ad-free — your support goes directly toward keeping it running.

or
Unlock all articles with Membership →
Share

Thank You for Reading

Gemini Lab is ad-free, supported entirely by members like you. We publish practical guides daily with implementation code, benchmarks, and production-ready patterns. If you've found it useful, we'd love to have you on board.

  • Copy-paste ready implementation code
  • New advanced guides published daily
  • $5/mo or $10 for lifetime access
View Membership →

Related Articles

API / SDK2026-05-03
Launching a Paid Service on Gemini API — A 2026 Roadmap
A practical 2026 roadmap for monetizing a service built on Gemini API — covering model selection, unit economics, pricing models, and the architectural decisions that decide whether your low API costs become a competitive edge or a price-war trap.
API / SDK2026-05-03
Gemini API Prepaid Billing Migration 2026 — Impact and Pre-Flight Checklist
Gemini API is moving to a prepaid billing model. Here's exactly what changes, what breaks if you ignore it, and the pre-flight checklist I used for my own production services.
API / SDK2026-04-26
From Free Tier to First Paying User with the Gemini API — Three Walls Indie Devs Hit
Reaching 'it works' with the Gemini API is easier than ever. Reaching 'someone paid for it' is a different problem entirely. Here are the three non-technical walls indie developers hit before their first paying user — with the minimal code that wires payment to key issuance.
📚RECOMMENDED BOOKS
Build a Large Language Model (From Scratch)
Sebastian Raschka
LLM Dev
Prompt Engineering for LLMs
Berryman & Ziegler
Prompting
AI Engineering
Chip Huyen
AI Eng
* Contains affiliate links
See all →