All Articles
    NewsPublished September 7, 20266 min read

    Google AI Mode Now Runs on Gemini 3.8 Flash

    Google added its newest workhorse model to AI Mode on launch day, the first time a Gemini release has arrived in Search on the same schedule. What it means for brands that already optimise for AI Mode, and why the model menu now matters to subscribers.

    Matiss Katanenko

    Matiss Katanenko

    Co-founder, Honeyb

    Google AI Mode Now Runs on Gemini 3.8 Flash

    On 2 September 2026 Google released Gemini 3.8 Flash and, for the first time, made it available in AI Mode the same day. The model now appears in the AI Mode model menu for Google AI Pro and Ultra subscribers, sitting between the default 'Auto' selection and the more expensive Pro tier, which positions AI Mode as a surface that receives new models on Google's launch schedule rather than following a day or two behind. For brands already optimising for AI Mode answers, the change raises a practical question: does 3.8 Flash name different brands, surface different sources, or behave materially differently from the 3.7 Flash model it replaces, and should you be testing your prompts against it immediately.

    The short answer is yes, because the model a subscriber chooses can reshape the answer, and because enough subscribers will switch to 3.8 Flash on announcement day that ignoring it until later in the month means missing a visibility shift while it is still reversible. This piece sets out what 3.8 Flash brings, why it landed in AI Mode faster than prior releases, and what the practical steps are for a brand that treats AI search as a measurable channel.

    What Gemini 3.8 Flash is built for

    Google describes 3.8 Flash as "our most intelligent workhorse model" with "significant improvements from 3.7 Flash across software engineering, agentic tasks, and critical, multi-step reasoning". The benchmarks Google published show it outperforming most larger frontier models on autonomous software engineering tasks (54.9% on HLE-Verified, strong performance on DeepSWE v1.1), and the model is designed to "work harder" by executing extra reasoning steps and calling tools iteratively, which can mean using more tokens on a complex task in exchange for a better result.

    For a brand optimising towards AI Mode, the salient detail is not the benchmark score but the reasoning behaviour, because a model that takes more steps to reach an answer may also consult more sources, weight different signals, or surface brands that a faster, shallower model skips. Whether that turns into a material change in which brands appear depends on the category and the query, which is why the first move after any model swap is to rerun your prompts and compare the results rather than assume continuity.

    The pricing remains the same as 3.7 Flash during the introductory period: $0.75 per million input tokens and $3.75 per million output tokens, rising to $1.50 and $7.50 after 31 December 2026. For most subscriber use the token cost is invisible, but for API-based monitoring of AI Mode at scale the December jump will matter.

    AI Mode now gets models on launch day

    The detail that signals a shift in how Google is treating AI Mode is not the model itself but the timing. Gemini 3.8 Flash appeared in the AI Mode model menu on 2 September, the same day Robby Stein, VP of Product for Google Search, announced it on X and the same day Google published the official blog post. Prior Gemini releases typically reached AI Mode a day or more after the wider launch, which positioned Search as a recipient of models rather than a launch surface, whereas this release treated AI Mode as part of the announcement itself.

    That changes the tempo for anyone tracking AI visibility across Google's surface, because a model that lands in subscriber hands on day zero can shift answers immediately rather than over a phased rollout, and because subscribers who follow Google AI announcements will switch to the new model within hours of it appearing in the menu. If your brand's visibility in AI Mode answers depends on which model is running, and you only check once a fortnight, you will miss the shift until it has already reshaped a week or two of answers.

    The model menu itself now lists 3.8 Flash under "Gemini 3 models", between the default 'Auto' selection and the Pro tier. The prior model, 3.7 Flash, no longer appears in the menu, which suggests Google is retiring it from subscriber-facing surfaces rather than keeping a prior-generation workhorse available. Free-tier users do not see the model menu at all and receive whichever model Google sets as the default, which has not been confirmed publicly but is likely 3.8 Flash given the menu structure.

    Want to get recommended by AI?

    Check your AI search visibility, then let the Honeyb agent write, fix, and earn what gets you recommended. Free to start.

    Free AI visibility checker

    What brands should test now

    If you already track your visibility in AI Mode, the move is to rerun your buyer-intent prompts against 3.8 Flash and diff the results against the baseline you have from 3.7 Flash or the default model. The questions to answer are whether the new model names your brand more or less often, whether it cites different sources, and whether the framing of the answer has shifted in a way that helps or harms your positioning. If the answer is materially worse, the fix is the same as it would be for any AI visibility drop: audit which sources the model is reading instead of yours, and get your brand into those sources through third-party mentions, community discussion, or structured content that is easier to lift.

    If you do not yet track AI Mode systematically, this is a forcing function to start, because models will continue to ship on this cadence and each one can reshuffle the answer. The baseline to set now is your recommend rate across your category's prompts with 3.8 Flash as the selected model, measured over at least a dozen runs per prompt to account for the variance AI answers carry. From there, any change in recommend rate after the next model release tells you whether the new model is better or worse for your brand, and which prompts moved.

    The other surface to watch is whether Google makes 3.8 Flash the default model for free-tier users, because that would shift the majority of AI Mode traffic onto the new model and make its behaviour the new baseline for everyone not actively selecting a different model. Google has not announced that change, but the menu structure and the retirement of 3.7 Flash suggest it is likely. If you are only tracking the default model and it quietly swaps to 3.8 Flash, you will see a visibility shift without an obvious cause unless you are logging which model was used for each check.

    For brands that already work with AI visibility tools, the question to ask your vendor is whether they track which model produced each answer and whether they can rerun historical prompts against 3.8 Flash so you can compare the two directly. If the tool only checks the default model and does not log which model that was, you will not be able to attribute a visibility change to the model swap rather than to a content or source shift elsewhere.

    Why this matters for AI search as a channel

    The broader implication is that AI Mode is now part of Google's model launch process, which means it will receive new Gemini models on announcement day rather than in a second wave, and that changes the tempo at which brands need to track and respond to shifts in AI-generated answers. If a model swap can change which brands appear and which sources are cited, and if that swap happens without warning on the day of a Google blog post, then tracking AI visibility once a month is no longer frequent enough to catch and respond to a drop before it has been live for weeks.

    That does not mean you need to check every day, but it does mean you should rerun your prompts whenever Google announces a new model, and it means the answer to whether a visibility drop was caused by a model change, a competitor earning new mentions, or a source going stale now requires knowing which model was running when the drop happened. The discipline that supports that is logging the model ID alongside the answer for every check, which most AI visibility tools do not yet do by default but will need to as model swaps become routine.

    The short version is that AI Mode just became a faster-moving surface, and the brands that treat it as a static ranking will miss shifts that the brands tracking it as a rate across models and reruns will catch and fix.

    Frequently asked questions

    When did Google add Gemini 3.8 Flash to AI Mode?

    Google added Gemini 3.8 Flash to AI Mode on 2 September 2026, the same day the model was officially announced. This was the first time a Gemini model reached AI Mode on launch day rather than following a day or two behind.

    Can I choose which model AI Mode uses?

    Google AI Pro and Ultra subscribers can select a model from the AI Mode menu, which now includes Gemini 3.8 Flash alongside the default 'Auto' selection and the Pro tier. Free-tier users do not see the model menu and receive whichever model Google sets as the default.

    Will Gemini 3.8 Flash change which brands appear in AI Mode answers?

    It depends on the category and the query. Gemini 3.8 Flash is designed to execute extra reasoning steps and call tools iteratively, which can mean consulting more sources or weighting signals differently than prior models. The only way to know if it affects your brand is to rerun your buyer-intent prompts and compare the results to a baseline from the prior model.

    What should I test after a new model is added to AI Mode?

    Rerun your buyer-intent prompts with the new model selected and compare the answers to your baseline from the prior model. Check whether your brand is named more or less often, whether the sources cited have changed, and whether the framing helps or harms your positioning. Track your recommend rate across multiple runs to account for variance.

    How does Gemini 3.8 Flash differ from 3.7 Flash?

    Google describes 3.8 Flash as having significant improvements over 3.7 Flash in software engineering, agentic tasks, and multi-step reasoning. It is designed to work harder by executing extra reasoning steps and calling tools iteratively, potentially using more tokens on complex tasks. The pricing remains the same during the introductory period: $0.75 per million input tokens and $3.75 per million output tokens until 31 December 2026.

    Should I track which model AI Mode is using for each check?

    Yes, especially now that new models reach AI Mode on launch day. If you do not log which model produced each answer, you will not be able to tell whether a visibility shift was caused by a model swap or by a content change elsewhere. Most AI visibility tools do not yet log the model ID by default, but it is becoming a necessary data point as model swaps become routine.

    Matiss Katanenko

    About the author

    Matiss Katanenko

    Co-founder, Honeyb

    My name is Matiss Katanenko and I co-founded Honeyb, the AI visibility platform that tracks how ChatGPT, Gemini, Claude, Perplexity and the other major AI engines talk about brands. Before Honeyb I ran SEO for fast-growing companies across the US and Europe, including one of America's 500 fastest-growing companies. The numbers I am proudest of: taking a site from zero to 200,000 monthly visitors in five months, and over $10M in client revenue attributed to organic search. I still run experiments across ten-plus of my own domains to test what actually works in SEO, programmatic SEO and AI search, and those experiments are what this blog reports on. My focus today is AI search visibility: how brands get retrieved, ranked and referenced by LLMs. I'm based in Riga, Latvia. In my free time I'm in the sauna, on a padel court, or behind a drum kit.

    Free to start

    Get recommended by AI search models.

    Run a free AI search visibility check, then let the Honeyb agent do the work that gets you into the answers.

    ChatGPTClaudeGeminiPerplexity