GEMINI LABJP
VIDEO — Agentic video understanding reached 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite on September 1. The model navigates the timeline itself rather than sampling frames at a fixed rateTOKENS — Because it pulls transcripts, frames, or audio only when it needs them, Google measures up to 88% fewer tokens on long-form contentSCOPE — It works across both the Interactions and GenerateContent APIs. If you have costed out long-video work before, the assumptions have movedMUSIC — Lyria 3.5 entered public preview on September 3, generating full-length songs at 44.1 kHz stereoCONTROL — Lyria 3.5 accepts text and image inputs, with better musical coherence, more natural vocals, and finer control over duration and structureROBOTICS — gemini-robotics-er-2-streaming-preview is tuned for real-time streaming over the Live API, with function calling that blocks on physical robot actionsVIDEO — Agentic video understanding reached 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite on September 1. The model navigates the timeline itself rather than sampling frames at a fixed rateTOKENS — Because it pulls transcripts, frames, or audio only when it needs them, Google measures up to 88% fewer tokens on long-form contentSCOPE — It works across both the Interactions and GenerateContent APIs. If you have costed out long-video work before, the assumptions have movedMUSIC — Lyria 3.5 entered public preview on September 3, generating full-length songs at 44.1 kHz stereoCONTROL — Lyria 3.5 accepts text and image inputs, with better musical coherence, more natural vocals, and finer control over duration and structureROBOTICS — gemini-robotics-er-2-streaming-preview is tuned for real-time streaming over the Live API, with function calling that blocks on physical robot actions
Articles/API / SDK
API / SDK/2026-08-16Intermediate

Moving to the Batch Tier Cancels Out the Gemini 3.7 Flash Price Increase Exactly

Introductory pricing for Gemini 3.7 Flash ends on December 31, 2026, and the rate doubles the next day. Here is which models are affected, which are not, and how to project your January bill from the usage you already have.

Gemini API234Pricing4Cost designModel selectionIndie development3

Premium Article

Gemini 3.7 Flash went generally available on August 13. The announcement mentioned introductory pricing through December 31, 2026, so I opened the pricing page to see how much of a discount that was.

The numbers were identical to Gemini 3.6 Flash.

$0.75 in, $3.75 out. Not a discount attached to the new model — the same deadline applies across the Flash line. And a little further down the page, the story shifted. This is not a discount that expires. It is a rate that doubles.

Some models double. Some do not. If you cross into January with that line blurred, the invoice will draw it for you.

The introductory rate is not exclusive to 3.7 Flash

Here are the published numbers. Checked on August 16, 2026 against the Gemini Developer API pricing page, in USD per 1M tokens, paid tier, Standard.

ModelInput (2026)Input (2027+)Output (2026)Output (2027+)
gemini-3.7-flash$0.75$1.50$3.75$7.50
gemini-3.6-flash$0.75$1.50$3.75$7.50
gemini-3.5-flash$1.50$1.50$9.00$9.00
gemini-3.5-flash-lite$0.30$0.30$2.50$2.50

3.7 Flash and 3.6 Flash match on both input and output. So the usual assumption — newer means pricier — does not hold here. There is no cost argument against moving from 3.6 to 3.7.

The row that surprised me was 3.5 Flash. An older generation, at twice the rate of 3.7 Flash, with no introductory pricing at all, so it stays at $1.50 / $9.00 into next year. Put that next to 3.7 Flash after the change ($1.50 / $7.50) and the input is level while the output is cheaper on 3.7. There is no financial reason to sit on 3.5 Flash, this year or next.

If you have been deferring the move to a newer model until things settle down, that deferral is not saving you anything.

Some models rise, some hold

Look again at the right half of that table. Only 3.7 Flash and 3.6 Flash move. 3.5 Flash and 3.5 Flash-Lite hold.

This is not an across-the-board increase. It changes the ratio between models.

Comparison20262027+
3.7 Flash ÷ 3.5 Flash-Lite (input)2.5×5.0×
3.7 Flash ÷ 3.5 Flash-Lite (output)1.5×3.0×

If you route between a light model and a heavy one, that shift lands directly on your design. A "when in doubt, escalate" policy that is affordable today may not be affordable in January. I will put numbers on the escalation rate later in this article, but the point to hold onto is that the increase acts on your architecture, not on a single model.

There is a second reading of the same fact. Because Flash-Lite holds, the incentive to push work down to it gets stronger next year. If quality allows, the next few months are a good window to widen the range of tasks Flash-Lite can handle.

Thank you for reading this far.

Continue Reading

What follows includes implementation code, benchmarks, and practical content we hope you'll find useful. This site runs without ads — server and development costs are supported entirely by members like you. If it's been helpful, we'd be truly grateful for your support.

WHAT YOU'LL LEARN
You will be able to project how much your Gemini spend rises on January 1, 2027 from your current usage, without waiting for the invoice
You will be able to tell which models are affected by the increase and which are not, so you are not hunting for a migration target in January
You will be able to separate the work that can wait from the work that cannot, and assign each to the Standard or Batch tier on your own criteria
Secure payment via Stripe · Cancel anytime

Unlock This Article

Get full access to the rest of this article. Buy once, read anytime. This site is ad-free — your support goes directly toward keeping it running.

or
Unlock all articles with Membership →
Share

Thank You for Reading

Gemini Lab is ad-free, supported entirely by members like you. We publish practical guides daily with implementation code, benchmarks, and production-ready patterns. If you've found it useful, we'd love to have you on board.

  • Copy-paste ready implementation code
  • New advanced guides published daily
  • $5/mo or $15 for lifetime access
View Membership →

Related Articles

API / SDK2026-05-03
Launching a Paid Service on Gemini API — A 2026 Roadmap
A practical 2026 roadmap for monetizing a service built on Gemini API — covering model selection, unit economics, pricing models, and the architectural decisions that decide whether your low API costs become a competitive edge or a price-war trap.
API / SDK2026-05-03
Gemini API Prepaid Billing Migration 2026 — Impact and Pre-Flight Checklist
Gemini API is moving to a prepaid billing model. Here's exactly what changes, what breaks if you ignore it, and the pre-flight checklist I used for my own production services.
API / SDK2026-04-26
From Free Tier to First Paying User with the Gemini API — Three Walls Indie Devs Hit
Reaching 'it works' with the Gemini API is easier than ever. Reaching 'someone paid for it' is a different problem entirely. Here are the three non-technical walls indie developers hit before their first paying user — with the minimal code that wires payment to key issuance.
📚RECOMMENDED BOOKS
Build a Large Language Model (From Scratch)
Sebastian Raschka
LLM Dev
Prompt Engineering for LLMs
Berryman & Ziegler
Prompting
AI Engineering
Chip Huyen
AI Eng
* Contains affiliate links