Obsurfable

Gemini 3.7 Flash Is Now Selectable in AI Mode — Re-Baseline Before It Becomes Default

Obsurfable

Google introduced Gemini 3.7 Flash on 13 August 2026 as its "most intelligent workhorse model yet for coding and agents" — three weeks after 3.6 Flash, at an introductory API price of $0.75 / $3.75 per million input/output tokens (half of 3.6 Flash's original rate through 31 December 2026).

The visibility-relevant event landed the next day. Search Engine Journal reported that Robby Stein, VP of Product for Google Search, said 3.7 Flash is now a selectable model in AI Mode, global in English for Google AI Pro and Ultra subscribers. Stein's line: it is "better at following instructions + understanding your intent."

It is not (yet) documented as AI Mode's default. It is also not on the free tier. That is exactly how previous Flash models entered Search — as a menu option — before they became the production path.

The Flash-to-default pattern

DateWhat happened in AI Mode
Nov 2025Gemini 3 Pro added as an option
Dec 2025Gemini 3 Flash became the default
May 2026 (I/O)Gemini 3.5 Flash became the global default
14 Aug 2026Gemini 3.7 Flash added as a selectable model (Pro/Ultra, English)

Jeff Dean has said Flash stays Search's production tier because of latency and cost. If that cycle continues, 3.7 Flash is not a developer sidebar. It is the candidate for the next default behind AI Mode — and, eventually, the surfaces that inherit the same routing.

Google has not said whether 3.7 Flash is in the Auto router that AI Mode started building in November 2025. Support docs still lag the live menu (Fast/Pro language versus Auto + new Flash). Until Google is explicit, assume Auto can change what "AI Mode" means without a press release.

What Google says the model is better at

From the 3.7 Flash announcement (vendor benchmarks, not citation studies):

  • Coding: FrontierCode 1.1 Main 43.6% vs 34.4%; DeepSWE v1.1 65.3% vs 49.0% versus 3.6 Flash
  • Documents / knowledge work: GDP.pdf 34.0% vs 22.0%
  • Business workflows: AutomationBench 30.4% vs 17.0%
  • Web UI generation: WebDev Arena Elo 1588 vs 1538
  • Agents: Gemini Spark (Pro/Ultra, 160+ countries) moved onto 3.7 Flash the same day, with tighter Google Workspace tool use

Instruction-following and intent understanding are the claims Stein attached to Search. For AEO, that usually shows up as:

  • Different query fan-out (which sub-questions get searched)
  • Different willingness to use tools and interactive layouts
  • Different citation mix when the model is more confident synthesizing vs quoting

We already documented a citation reset when Gemini 3 landed: top rankings no longer predicted AI Overviews. 3.5 Flash as default was another routing change. 3.7 Flash is the next candidate. Do not wait for the default flip to discover it.

How to use the model menu as a measurement instrument

For every priority prompt, if you have Pro/Ultra:

  1. Run Auto (what most subscribers actually get)
  2. Run 3.7 Flash explicitly
  3. Run Pro if it remains in the menu

Compare mention, citation URLs, and competitor substitution. If Flash and Auto already diverge, Auto may already be routing some queries to 3.7 — or 3.7 is a preview of the next default.

Log:

  • Date and locale
  • Model selected (do not write "AI Mode" as if it were one model)
  • Whether generative UI / tools appeared in the answer (that surface is expanding in parallel)

Free-tier and logged-out AI Mode remain a separate column. Most of the world is not on Ultra.

What not to do

  • Do not rebuild your content program around coding benchmarks. Those explain why Google shipped the model. They do not tell you who gets cited for "best CRM for startups."
  • Do not treat a 3.7 Flash demo as the default AI Overview. Overviews and AI Mode are still distinct surfaces with historically limited URL overlap.
  • Do not skip the re-baseline because "it's only a selector." Selectors become defaults on Google's calendar, not yours.

Pricing tells you Google wants this in production

Introductory API pricing of $0.75 / $3.75 per million tokens through year-end (then $1.50 / $7.50) is not a research-preview tariff. Combined with Spark moving onto 3.7 Flash for Pro/Ultra in 160+ countries, Google is putting the model where agents already run 24/7 — email drafts, Workspace tools, multi-step jobs.

Search is latency-sensitive. Flash remains the plausible default because it is cheap and fast. Pro stays the "think harder" lane. Your citation program should mirror that: a Fast/Flash column and a Pro column, not one "Gemini" blob.

What 3.5 taught us (do not repeat the miss)

When 3.5 Flash became the AI Mode default at I/O, teams that had baselined 3 Pro spent June explaining "random" citation churn. The churn was routing. Document 3.7 now, while it is still a selector, so the next default is a labeled discontinuity on the chart.

Related: What Gemini 3.5 Changes About AI Search Visibility.

How Obsurfable fits

When the model behind a Google AI surface changes, Obsurfable is the control plane: same prompts, dated runs, engine-separated tables. The Visibility Director should treat a Flash-selector week as a model-change protocol — the same 7-day re-baseline you use for ChatGPT defaults.

Bottom line

Gemini 3.7 Flash is in AI Mode for paying subscribers now, as a workhorse model Google is pricing to run at scale. History says Flash does not stay optional. Measure Auto vs 3.7 vs Pro this week so the next default does not look like a mysterious citation crash.