Gemini 3.8 Flash Pricing: What You Pay and When
The introductory rate, the 2027 regular rate, app access, and a worked cost example
Gemini 3.8 Flash pricing comes in two stages: a low introductory rate now, followed by a higher regular rate. Google announced the model on September 2, 2026 (Google).
Through December 31, 2026, the price is $0.75 per 1M input tokens and $3.75 per 1M output tokens (Google). Starting January 1, 2027, those rates double.
This guide covers both rates, walks through a cost example, and explains how app users can get access. Always check Google's pricing page for the latest numbers.
How Much Does Gemini 3.8 Flash Cost?
Gemini 3.8 Flash costs $0.75 per 1M input tokens and $3.75 per 1M output tokens during its introductory window (Google). That window runs through December 31, 2026.
From January 1, 2027, the regular rate applies: $1.50 per 1M input tokens and $7.50 per 1M output tokens (Google). That is double the intro rate.
These are Google's published API rates. Prices can change without notice, so verify current numbers on Google's pricing page before you commit.
| Rate | Input (per 1M tokens) | Output (per 1M tokens) | When |
|---|---|---|---|
| Introductory | $0.75 | $3.75 | Through Dec 31, 2026 |
| Regular | $1.50 | $7.50 | From Jan 1, 2027 |
Run Your AI On Mac Studio

The ultimate machine for running AI models on your own desk: M5 Max, a 32-core GPU, and 36GB of unified memory.
Why the Rate Matches Gemini 3.7 Flash
Gemini 3.8 Flash launches at the same API rate as Gemini 3.7 Flash (Google). Both charge $0.75 input and $3.75 output per 1M tokens during the intro window, and both rise to $1.50 and $7.50 on January 1, 2027 (Google).
So this is a newer model at the same price, not a price cut. The value shift is in capability, since the price holds while Google claims gains over 3.7 Flash on reasoning and coding (Google).
The intro rate is half the regular rate (Google). If you already run 3.7 Flash, moving to 3.8 Flash does not change your per-token bill, which makes a pilot low-risk on cost alone.
A Worked Example: What 1M in and 1M out Costs
One million input tokens plus one million output tokens costs $4.50 at the intro rate (Google). That is $0.75 for input and $3.75 for output.
At the regular 2027 rate, the same job costs $9.00 (Google). That is $1.50 for input and $7.50 for output.
Output tokens cost five times as much as input tokens at both rates. So long answers, code, and reports drive most of your bill.
When we ship model-launch page families across our portfolio at Layer3 Labs, buyers ask about price before benchmarks. A simple input-plus-output example like this settles it faster than a spec sheet.
| Workload | Intro cost | Regular cost |
|---|---|---|
| 1M in + 1M out | $4.50 | $9.00 |
| 500K in + 500K out | $2.25 | $4.50 |
| 2M in + 1M out | $5.25 | $10.50 |
What Is a Token, and Why It Drives Your Bill
A token is a small chunk of text, roughly a few characters or part of a word. Gemini 3.8 Flash pricing is billed by the token, split into input and output (Google).
Input tokens are what you send: your question, any pasted text, and instructions. Output tokens are what the model writes back.
This matters because output is the expensive half. At both rates, output tokens cost five times as much as input tokens (Google). Short questions with long answers cost more than they look.
Watch this closely with Gemini 3.8 Flash. Google says the model works harder on complex tasks, taking extra reasoning steps and tool calls before it answers (Google). Those extra steps can add output tokens, so long agent runs may cost more than a simple prompt suggests.
Getting Gemini 3.8 Flash in the App
App users reach Gemini 3.8 Flash in the Gemini app on a Google AI Pro or Ultra plan (Google). This is a subscription path, not a per-token bill.
You pay a flat monthly fee for the plan, and the app gives you the model. Google also surfaces the model in AI Mode in Google Search and in Google Sheets, which your existing Google plan may already cover (Google).
Google's announcement did not detail a free consumer tier for the model (Google). Confirm current plan prices and any free access with Google before you subscribe.
Which Pricing Path Fits You?
Choose the subscription path if you want Gemini 3.8 Flash inside the app for daily chat and writing. A Google AI Pro or Ultra plan covers that (Google).
Choose the API path if you build software and pay per token. That is where the $0.75/$3.75 intro rate and $1.50/$7.50 regular rate apply (Google).
Most individuals fit the subscription. Teams building products or automations fit the API. Occasional writing and research point to the subscription, and steady programmatic calls point to the per-token API.
Some buyers use both. You might chat in the app on a Pro plan while your product team runs the API for automated tasks. The two bills are separate and do not overlap.
One case does not fit either path cleanly. If your workload depends on a fixed context or throughput limit, note that Google did not publish those figures at launch (Google), so confirm them before you pick a plan.
How to Verify Gemini 3.8 Flash Pricing
Verify Gemini 3.8 Flash pricing on Google's official pricing page before you buy or budget. Prices change without notice, and the intro rate ends December 31, 2026 (Google).
Check two things: the current per-token API rate and the current app subscription plans. Both live on Google's own pages, not on aggregators.
Set a reminder for late December 2026. That is when the intro rate expires and the regular rate begins (Google).
Do not budget off this page or any other third party. Rates and plans can move between now and your purchase, so treat Google's own page as the final word.
How to use Gemini 3.8 Flash
You do not host Gemini 3.8 Flash yourself — you use it through a tool, so "getting started" really means choosing the right one.
The fastest way to put Gemini 3.8 Flash to work day to day is inside an AI IDE, and Cursor is the most popular — it supports it directly, so you can be working in minutes. The maker's own option is Antigravity for Gemini 3.8 Flash, if you want the native experience. Prefer a different editor? Windsurf, Zed, and GitHub Copilot drive these models too.
Frequently Asked Questions
- Gemini 3.8 Flash costs $0.75 per 1M input tokens and $3.75 per 1M output tokens during the introductory window through December 31, 2026 (Google). From January 1, 2027, the regular rate is $1.50 input and $7.50 output per 1M tokens (Google). Verify current prices on Google's pricing page.
- The introductory price ends December 31, 2026 (Google). Starting January 1, 2027, the regular rate of $1.50 input and $7.50 output per 1M tokens applies (Google). The intro rate is half of the regular rate.
- No, they cost the same. Gemini 3.8 Flash launches at the same API rate as Gemini 3.7 Flash: $0.75 input and $3.75 output per 1M tokens during the intro window, then $1.50 and $7.50 from January 1, 2027 (Google). The change is in capability, since the price is the same.
- You reach Gemini 3.8 Flash in the Gemini app on a Google AI Pro or Ultra plan (Google). Google also surfaces the model in AI Mode in Google Search and in Google Sheets. Confirm current plans with Google.
- Google's announcement did not detail a free consumer tier for Gemini 3.8 Flash (Google). App access is tied to a Google AI Pro or Ultra plan. Confirm current free-tier availability with Google before assuming any free quota.
- One million input tokens plus one million output tokens costs $4.50 at the intro rate and $9.00 at the regular rate (Google). Output tokens cost five times as much as input tokens, so long answers drive most of your bill.
Not Sure Which Gemini 3.8 Flash Plan Fits?
Book a free 30-minute AI workflow audit with Layer3 Labs. We will map Gemini 3.8 Flash pricing to your real usage and tell you whether the subscription or the API path saves you more.
Book an Audit