Skip to main content
LLMgram · AI News · 2026-09-03

Google Ships Its Third Gemini Flash Model in Six Weeks

Google Ships Its Third Gemini Flash Model in Six Weeks

Google DeepMind released Gemini 3.8 Flash and Gemini 3.8 Flash Cyber on September 2, 2026, marking a third Flash-tier rollout in six weeks while Gemini Pro updates remain absent. Both variants share one foundational intelligence divided by safety mitigations rather than model size. Gemini 3.8 Flash is generally available at introductory rates of $0.75 and $3.75 per million input and output tokens through December 31, 2026. Google says it beats Claude Opus 5 and GPT-5.6 Sol on some benchmarks; reporting adds agentic coding parity with Opus 5 at lower list pricing. Flash Cyber, aimed at Fairwind Program partners, reached 47.2% pass@1 on CWE-Bench. Reporting cautions that heavier reasoning can consume roughly thirty percent more output tokens per task than its predecessor, so real bills may outrun headline token rates.

Sources

Google Ships Its Third Gemini Flash Model in Six Weeks

Google Ships Its Third Gemini Flash Model in Six Weeks

Google has not released a Gemini Pro model in a while, and on Wednesday launched yet another Flash-tier set. The New Stack counts this as Google's third Gemini Flash shipment in six weeks.

Key takeaway

Google is accelerating Flash-tier Gemini releases while Pro-tier frontier updates stay missing, pairing benchmark claims with separate general and cyber access envelopes.

What happened

On September 2, 2026, Google DeepMind released Gemini 3.8 Flash and Gemini 3.8 Flash Cyber. MarkTechPost reports both variants run on the same foundational intelligence, split by safety mitigations rather than model size, with Gemini 3.8 Flash generally available at $0.75 and $3.75 per 1M tokens through December 31, 2026.

Coverage counts this as Google's third Flash shipment in six weeks, arriving three weeks after Gemini 3.7 Flash. Google says Gemini 3.8 Flash beats Claude Opus 5 and GPT-5.6 Sol on some benchmarks and launched Flash Cyber for Fairwind Program partners; The Decoder notes higher output-token use on some reasoning tasks despite identical token rates.

Evidence

  • Google released Gemini 3.8 Flash and Gemini 3.8 Flash Cyber on September 2, 2026.

    MarkTechPost · attributed

    Google released Gemini 3.8 Flash and Gemini 3.8 Flash Cyber on September 2, 2026.

  • Both variants share one foundational intelligence split by safety mitigations rather than model size.

    MarkTechPost · attributed

    Both variants run on the same foundational intelligence, split by safety mitigations rather than model size.

  • Gemini 3.8 Flash introductory pricing is $0.75 input and $3.75 output per 1M tokens through December 31, 2026.

    Techmeme · attributed

    Google releases Gemini 3.8 Flash, three weeks after Gemini 3.7 Flash, for an introductory price of $0.75/1M input and $3.75/1M output tokens until December 31

  • Google says Gemini 3.8 Flash beats Claude Opus 5 and GPT-5.6 Sol on some benchmarks.

    Techmeme · attributed

    Google launches Gemini 3.8 Flash Cyber for partners in its new Fairwind Program and says Gemini 3.8 Flash beats Claude Opus 5 and GPT-5.6 Sol on some benchmarks

  • Flash Cyber reached 47.2% pass@1 on CWE-Bench.

    MarkTechPost · attributed

    Flash Cyber reaches 47.2% pass@1 on CWE-Bench

  • The Decoder reports Gemini 3.8 Flash reasoning can burn about 30 percent more output tokens per task than its predecessor.

    The Decoder · attributed

    its "working harder" reasoning burns about 30 percent more output tokens per task, making it pricier in practice than its predecessor despite identical token rates

  • The New Stack counts this as Google's third Gemini Flash shipment in six weeks.

    The New Stack AI · attributed

    The New Stack counts this as Google's third Gemini Flash shipment in six weeks.

Why it matters

The cadence signals Google is prioritizing fast, lower-cost Flash iterations for production workloads while a new Gemini Pro frontier release remains absent from recent reporting.

Limits and uncertainties

Benchmark wins cited by Google are described as on some benchmarks only, not across all tasks or workloads.

The Decoder notes higher output-token consumption on some reasoning tasks, so list token rates may understate real spend.

Reporting does not establish when a new Gemini Pro-tier model will ship.

Practical implications

Evaluate Gemini 3.8 Flash on agentic coding and cybersecurity workloads where Flash Cyber and Fairwind Program access apply.

Budget using measured output-token volume through December 31, 2026 introductory pricing rather than headline per-million rates alone.

Track whether repeated Flash releases change default routing away from prior Gemini 3.7 Flash deployments.

What to watch

Whether Google announces a Gemini Pro-tier frontier model after this third Flash release in six weeks.

Post-introductory pricing after the December 31, 2026 deadline for Gemini 3.8 Flash.

Independent cost benchmarks comparing total tokens per task against Gemini 3.7 Flash and Claude Opus 5.

Sources

LLMgram editorial selection and synthesis · @llmgram. LLMgram is not the original publisher of this information.
Continue on LLMgram: Open in AI Signal →
Original reporting: Google ships its third Gemini Flash model in six weeks