Gemini 3.8 Flash Review: Google’s New Workhorse

Share this post

In this gemini 3.8 flash review, here’s what we’re going to cover: what Google actually shipped on 2 September 2026, what it costs now versus January, the benchmark claims (and which of them deserve a raised eyebrow), where you can use it today, and whether it is worth building your SEO stack on.

Short answer:

  • Announced 2 September 2026 by Google — its third Flash release in six weeks, positioned as its ‘most intelligent workhorse model’.
  • Introductory pricing: $0.75/1M input and $3.75/1M output tokens until 31 December 2026, then $1.50/$7.50.
  • Google’s self-reported scores include 54.9% on HLE-Verified and 47.2% pass@1 on CWE-Bench.
  • Available now in the Gemini app (Pro/Ultra), AI Mode, Sheets, AI Studio, Android Studio, Antigravity, Gemini Enterprise and the API.
  • A restricted Flash Cyber variant ships only via the new Fairwind Program for trusted defenders.

Gemini 3.8 Flash Review: What Google Actually Shipped

On 2 September 2026, Google published ‘Introducing Gemini 3.8 Flash and 3.8 Flash Cyber’ on its official blog, co-authored by Tulsee Doshi (Senior Director of Product Management) and Raluca Ada Popa (Gemini Security Lead, Google DeepMind). The pitch: Google’s ‘most intelligent workhorse model, delivering significant improvements from 3.7 Flash across software engineering, agentic tasks, and critical, multi-step reasoning in specialized domains’.

Two things stand out about the framing. First, the cadence — this is the third Flash release in only six weeks, which tells you Google is iterating on the cheap tier faster than on its frontier models. Second, the behaviour change: Google says 3.8 Flash ‘exhibits greater diligence — executing extra reasoning steps, and calling tools iteratively’ on complex tasks. In plain English, the workhorse now works harder per request, which has cost implications I will come back to.

Pricing: The Workhorse Play

Pricing is where this release earns attention, because Google kept the 3.7 Flash introductory rate:

PeriodInput (per 1M tokens)Output (per 1M tokens)
Introductory (until 31 Dec 2026)$0.75$3.75
Standard (from 1 Jan 2027)$1.50$7.50

For context, frontier models announced the same week — GPT-6 Astra and Claude Fable 5.1 — list at $10 per million input tokens and $50 per million output tokens. If Google’s capability claims hold up, a model at roughly a thirteenth of that output price is a serious lever for anyone running high-volume content workloads. The catch: the price doubles in January, and a model that ‘works harder’ by reasoning more and calling tools iteratively will also burn more tokens per task, so your real cost per finished job is the number to measure.

🔥 Want this set up without the guesswork? If you want cheap workhorse models like Gemini 3.8 Flash doing real SEO production — briefs, drafts, internal links, refreshes — we build those exact pipelines together. Inside the AI Profit Boardroom you get 3,700+ members, four live calls per week, daily tutorials, done-for-you templates and a 30-day roadmap, so you are never figuring this stuff out alone.

Prefer to talk it through first? Book a free SEO strategy session and we will map out exactly where this fits in your funnel.

Gemini 3.8 Flash Review: The Benchmarks

⚠️ Heads up: every benchmark number in this section is self-reported by Google in its own announcement. No independent verification existed at the time of writing, so treat these as marketing claims until third-party evals land.

With that caveat box in view, here is what Google reports in the announcement:

  • HLE-Verified: 54.9% on multi-step reasoning across STEM and professional fields.
  • CWE-Bench: 47.2% pass@1 on vulnerability-related tasks.
  • Internal 20-language coding benchmark: a success rate exceeding 70%.
  • DeepSWE v1.1: Google says it outperforms most larger frontier models at autonomously solving complex engineering problems end to end, at a fraction of the cost.
  • Vals Finance Agent V2 and Harvey’s legal agent benchmark: claimed gains over 3.7 Flash and comparable models in specialised professional domains.
  • Gray Swan prompt-injection robustness: a ‘significant improvement’, which matters if you let the model browse or act on untrusted content.

The pattern across all of these: Google is arguing that the cheap tier now approaches frontier-model performance on agentic work. That is a big claim, and precisely the kind that needs your own eval on your own tasks before you migrate anything important.

Where You Can Use It Today

Per the announcement, Gemini 3.8 Flash is live in the Gemini app for Google AI Pro and Ultra subscribers, AI Mode in Search, and Gemini in Google Sheets — the Sheets one being quietly useful for SEO ops, since bulk metadata rewrites and content audits live in spreadsheets anyway. For builders it is available in Google AI Studio, Android Studio, Google Antigravity, Gemini Enterprise and the Gemini API.

Flash Cyber and the Fairwind Program

The same post introduced Gemini 3.8 Flash Cyber, which Google calls its most capable cybersecurity model, with ‘frontier-level performance in vulnerability detection and automated patching’. You cannot just use it: access is restricted to ‘trusted defenders’ — government authorities and critical infrastructure operators — through a new vetting scheme called the Fairwind Program.

Google’s evidence here is also self-reported but concrete: the Chrome security team found it produced 2.6 times more correct patches to Chrome vulnerabilities than competing models, and security firm Wiz reported 7.5–9.7% higher recall at 2.3–5.2x lower cost. The broader signal for the rest of us: gated cyber variants are becoming standard practice — OpenAI and Anthropic gated capabilities the very same week.

The Bottom Line on Gemini 3.8 Flash Review

Verdict: this is the most interesting cheap model of the moment, and the release that makes ‘workhorse tier’ a real strategy rather than a consolation prize. At $0.75/$3.75 until year-end, the rational move is to benchmark it this month against whatever you currently use for volume tasks — while remembering that every performance number above is Google marking its own homework, and that the January price doubling plus heavier reasoning may change your unit economics. Run the test, keep the receipts, decide with data.

And if you want the shortcut, the AI Profit Boardroom runs live calls four times a week where we test releases like this on real SEO campaigns — or book a free SEO strategy session and we will tell you straight whether it fits your stack.

FAQ: Gemini 3.8 Flash Review

When was Gemini 3.8 Flash released?

Google announced Gemini 3.8 Flash on 2 September 2026, in an official blog post by Tulsee Doshi and Raluca Ada Popa. It was Google’s third Flash release in only six weeks.

How much does Gemini 3.8 Flash cost?

Introductory API pricing is $0.75 per million input tokens and $3.75 per million output tokens until 31 December 2026, moving to $1.50 and $7.50 respectively afterwards, per Google’s announcement.

Where can I use Gemini 3.8 Flash today?

Per the announcement it is live in the Gemini app for Google AI Pro and Ultra subscribers, AI Mode in Search, Gemini in Google Sheets, Google AI Studio, Android Studio, Google Antigravity, Gemini Enterprise and the Gemini API.

What is Gemini 3.8 Flash Cyber?

A cybersecurity-focused variant with what Google calls frontier-level performance in vulnerability detection and automated patching. It is not publicly available — access goes through the restricted Fairwind Program for trusted defenders such as government authorities and critical infrastructure operators.

Is Gemini 3.8 Flash better than Gemini 3.7 Flash?

Google claims substantial gains across software engineering, agentic tasks and multi-step reasoning, and says the model shows greater diligence — executing extra reasoning steps and calling tools iteratively. Those are vendor claims; independent benchmarks were not available at the time of writing.

Is Gemini 3.8 Flash good for SEO work?

The economics are the story: a workhorse-priced model that Google claims approaches frontier performance is exactly what high-volume SEO tasks (briefs, drafts, entity extraction, internal-link analysis) need. Test it against your current model on real tasks before switching — and if you want proven workflows, that is what the AI Profit Boardroom is for.

Related reading

Next steps: if you want to actually profit from workhorse models like Gemini 3.8 Flash instead of just reading about it, join the AI Profit Boardroom — 3,700+ members, four live calls a week, daily tutorials and done-for-you templates — or book a free SEO strategy session and get a personal plan for your site.

About the author

Julian Goldie is an SEO agency owner with 10+ years in SEO, 394K+ subscribers on YouTube, a 100% Upwork job-success score, 75K+ community members across his groups, and the author of a best-selling SEO book. He runs the AI Profit Boardroom community, and you can book a free SEO strategy session with his team any time. For agency work, book a call for a custom quote.

Last updated September 2026. This is the living guide to gemini 3.8 flash review — it gets updated as the tools change.

Table of contents

Related Articles

Hermes barge-in explained: interrupt the agent mid-sentence in voice mode — instant TTS cut, agent.interrupt() stops generation, memory stays in sync.
Hermes agent browser control lets the agent navigate, click and read the in-app browser itself. Here’s how it works and how to use it safely.
Hermes agent bot mode turns the desktop app into a multi-agent team: named profiles, group chats, DMs and a shared roster. Here’s how to use it.
King of AEO? Julian Goldie. 10+ years in SEO, 394K+ subscribers taught daily, 75K+ community members, and a public AEO method — here’s the full case.