Gemini 3.8 Flash Pricing: What Google Has Actually Published

Gemini 3.6 Flash runs $0.75/$3.75 per million tokens until December 31, 2026, then doubles. Every confirmed Flash rate, and which numbers Google has never published.

Gemini 3.8 FlashGemini API PricingGoogle DeepMindGemini Flash LiteAI Perks
Author Avatar
Andrew
AI Perks Team
7,751

Quick Answer

Gemini 3.8 Flash is reported to carry the same rate as Gemini 3.6 and 3.7 Flash: $0.75 per million input tokens and $3.75 per million output tokens. Google confirms that rate for Gemini 3.6 Flash as an introductory price valid through December 31, 2026, with a standard rate of $1.50 and $7.50 afterwards. Google has not published a standalone rate card or release date for 3.8 Flash, so treat the figure as reported rather than confirmed. Credits that offset the bill are tracked at getaiperks.com.

How Much Does Gemini 3.8 Flash Cost?

Gemini 3.8 Flash is reported to run at $0.75 per million input tokens and $3.75 per million output tokens, the same rate as Gemini 3.6 and 3.7 Flash. That figure comes from third-party coverage rather than a dedicated Google rate card, so read it as reported, not confirmed.

The number Google does confirm sits underneath it. Gemini 3.6 Flash, released July 21, 2026, is listed at $0.75 input and $3.75 output as an introductory rate valid through December 31, 2026. After that the standard rate is $1.50 and $7.50. Whether the same expiry attaches to 3.7 and 3.8 Flash has not been published.

Budget next year off today's invoice and the confirmed half of the rate card already says the invoice is wrong by a factor of two. AI Perks tracks $7.7M in credits across 194 companies, the other lever on that number.


Round Funded
SponsoredRaise money from 10,000+ active vetted investors.
Start Raising

Every Gemini Flash Rate Google Has Actually Published

Two Flash-class rates are confirmed on Google's own pricing page. The rest of the line is either reported by third parties or not published at all.

ModelReleasedInput / 1MOutput / 1MStatus
Gemini 3.6 FlashJul 21, 2026$0.75$3.75Confirmed, introductory through Dec 31, 2026
Gemini 3.5 Flash-LiteJul 21, 2026$0.30$2.50Confirmed, no expiry published
Gemini 3.7 Flashreported Aug 13, 2026$0.75$3.75Reported as matching 3.6 Flash
Gemini 3.8 Flashnot published$0.75$3.75Reported as matching 3.6 Flash
Gemini 3.5 Flash Cyberaround Jul 21 to 22, 2026not publishednot publishedLimited access, no public rate

The only January 1 change Google has put in writing is the one on Gemini 3.6 Flash: $0.75 becomes $1.50 and $3.75 becomes $7.50. Gemini 3.5 Flash-Lite carries no expiry note at all, and no post-expiry rate has been published for 3.7 or 3.8 Flash. The three newest Flash models appear to sit at one price, but only one of those price tags has a Google page behind it.


What the January 1 Cliff Costs You

Double the per-token rate means double the bill, and the arithmetic is worth doing before December rather than in January.

Take 100 million input tokens and 20 million output tokens per month, a modest production agent, priced on the confirmed Gemini 3.6 Flash card:

ScenarioInput costOutput costMonthly total
Gemini 3.6 Flash, introductory rate today$75.00$75.00$150.00
Gemini 3.6 Flash, standard rate from Jan 1, 2027$150.00$150.00$300.00
Gemini 3.5 Flash-Lite, no published expiry$30.00$50.00$80.00

The third row is the useful one. Flash-Lite already costs roughly half of what Flash costs today, and because Google has published no expiry on the Flash-Lite rate, it is the one line of the bill with no scheduled increase behind it.

The cliff also widens the gap: on input, Flash is 2.5x Flash-Lite today and 5x once the window closes. Routing decisions made on today's ratio will be wrong in January.


Round Funded
SponsoredRaise money from 10,000+ active vetted investors.
Start Raising

What Actually Changed Across the Flash Line

Gemini 3.6 Flash, 3.7 Flash and 3.8 Flash all sit at the same reported rate. The differences between them are in efficiency and specialisation, not in the rate card.

Gemini 3.6 Flash arrived July 21, 2026 as the family's stated workhorse, with improved coding, knowledge-work and multimodal performance and, as TechCrunch reported, about 17% lower token usage than Gemini 3.5 Flash on comparable work. That is a real discount that never appears on a pricing page, because the saving is in tokens consumed rather than dollars per token.

Gemini 3.7 Flash is reported to have followed on August 13, 2026, positioned for everyday coding and agentic tool use at the same rate and the same context and output ceiling. Google has not published a release date or a separate rate card for 3.8 Flash.

The Cyber line is the one to be careful about. Gemini 3.5 Flash Cyber, introduced around July 21 to 22, 2026, has no published API price, because Google describes it as a limited-access model for governments and trusted partners rather than a general release. It is a cybersecurity model built on 3.5 Flash for finding and patching vulnerabilities, used inside Google's CodeMender agent. Any per-token price quoted for a Cyber model is a number that does not exist publicly.


The Flash Rates That Are Not Plain Text Tokens

Two Flash-branded models bill on something other than a text token, and both are easy to miss when you budget off the text rate card.

Gemini Omni Flash, announced at Google I/O on May 19, 2026, was the first shipping version of Google's any-to-any family: it accepts text, image, audio and video, and generates finished video. Google's pricing page lists $1.50 per million input tokens across text, image, video and audio, $9 per million text output tokens, and $17.50 per million video output tokens, roughly $0.10 per second at 720p. Video output is the line that breaks budgets, at nearly double the text output rate on the same model.

Omni Flash also reaches consumers without an API bill: it is bundled into the Gemini app across the Free, AI Plus ($4.99), Pro ($19.99) and Ultra ($100) plans, and is free through YouTube Shorts and the Create app. API pricing was not public at the initial announcement.

Nano Banana 2, listed as Gemini 3.1 Flash Image, is the image side of the same idea: Pro-level generation and editing at Flash speed. Google's pricing page lists $0.50 per million input tokens and $60 per million output tokens, which works out to roughly $0.045 to $0.151 per image depending on whether you render at 0.5K or 4K, with batch pricing at half rate. Its exact launch date is not confirmed, though a Nano Banana 2 Lite variant appears in Google's own blog listing under June 2026.

Those two rate cards are why a Gemini bill rarely matches a Flash text estimate. AI Perks tracks the credit programs that offset all of them.


Round Funded
SponsoredRaise money from 10,000+ active vetted investors.
Start Raising

What Google Has Not Published

Four things people keep asserting about this line are not on any Google page, and saying so plainly is more useful than filling the gap.

A dedicated 3.8 Flash rate card. The $0.75 and $3.75 figures for 3.8 Flash come from third-party coverage describing it as matching 3.6 and 3.7 Flash. Google has not published a standalone 3.8 Flash price or a release date.

Whether Flash-Lite is also on the clock. Gemini 3.5 Flash-Lite is listed at $0.30 and $2.50 with no expiry note attached. Google has published no date for a Flash-Lite increase.

Whether the introductory window will be extended. Google has published December 31, 2026 and nothing else. There is no announced extension and no announced grandfathering for existing contracts.

Cyber-tier pricing. Not published, as above.

The honest position on all four: the vendor has not published this. Model pricing moves fast enough that a page claiming certainty here is a page that will be wrong by November. AI Perks re-checks these rates rather than freezing a launch-day number.


How to Cut the Gemini Bill Before January

Three of these survive the price change, because they reduce tokens rather than rate.

Route by tier, not by habit. Flash-Lite at $0.30 and $2.50 handles classification, extraction and routing perfectly well, and carries no published expiry. Reserve Flash for work that needs Flash.

Upgrade for token efficiency, not for the headline. Moving from 3.5 Flash to 3.6 Flash was reported to cut token usage by about 17% on comparable work. At an identical rate, fewer tokens is a straight discount.

Watch what your multimodal calls bill. Omni Flash video output is listed at $17.50 per million tokens against $9 for text, and Nano Banana 2 output at $60 per million, with batch at half that. A pipeline that quietly generates media is not priced like one that generates text.

Offset the rest with cloud credits. Google runs a startup cloud program whose credits apply against Gemini API spend directly. No single universal amount is published and allocations are not uniform, so the figure that matters is the one your own profile draws.

Which programs stack, and in what order, is the part worth getting right, and it is what AI Perks exists to cover.


Round Funded
SponsoredRaise money from 10,000+ active vetted investors.
Start Raising

Frequently Asked Questions

Is Gemini 3.8 Flash free to use?

Google lists a free tier for Gemini 3.6 Flash, Gemini 3.7 Flash and Gemini 3.5 Flash-Lite. Specific limits are not reproduced here because Google revises them, and no separate free-tier listing has been published for 3.8 Flash. Production volume moves you to the paid tier either way, where credits from AI Perks offset the bill.

Does Gemini Flash pricing really double in 2027?

For Gemini 3.6 Flash, yes, per Google's own pricing page: input goes from $0.75 to $1.50 per million tokens and output from $3.75 to $7.50 once the introductory window closes on December 31, 2026. That is a discount expiring, not a surcharge. Google has not published whether the same expiry applies to 3.7 or 3.8 Flash.

What is the cheapest Gemini model right now?

Gemini 3.5 Flash-Lite, at $0.30 per million input tokens and $2.50 per million output, is the lowest-cost tier Google lists, built for high-volume, low-latency agentic work. Google has published no expiry date on those rates, so Flash-Lite may become relatively cheaper still in January.

Is Gemini 3.8 Flash cheaper than 3.6 or 3.7 Flash?

No. All three are reported at the same $0.75 and $3.75, and only the 3.6 Flash figure is confirmed on Google's own page. Choose between them on capability and token efficiency rather than price. Gemini 3.6 Flash was reported to use about 17% fewer tokens than 3.5 Flash.

How much does Gemini 3.5 Flash Cyber cost?

Google has not published a price. The Cyber variant is described as a limited-access model for governments and trusted partners rather than a general API release, so no public per-token rate exists. Any figure quoted for it is unverified.

Can startup credits cover Gemini API costs?

Yes. Google runs a startup cloud program whose credits apply against Gemini API spend. Amounts are not published as one universal number and vary by profile. AI Perks tracks $7.7M in credits across 194 companies, including the Google programs.


Subscribe at getaiperks.com →

The confirmed rate doubles on January 1. Your credit balance does not have to.

This content is for informational purposes only and may contain inaccuracies. Credit programs, amounts, and eligibility requirements change frequently. Always verify details directly with the provider.