Google releases Gemini 3.8 Flash, its third new AI model in six weeks

A standard version for coding and everyday tasks, plus a security-focused variant, both priced at half the regular rate until the end of 2025.

AI2Day Newsdesk3 min read
Photoreal editorial shot of a sleek modern data centre corridor at night, rows of server racks glowing with soft blue and amber light, faint reflections on poli
Share

Key points

  • Google launched Gemini 3.8 Flash on the same day it became its third Flash model release in six weeks.
  • The model ships in two versions: a general-purpose "workhorse" and a cybersecurity-tuned edition called Gemini 3.8 Flash Cyber.
  • Introductory pricing is $0.75 per million input tokens (small chunks of text fed to the model) and $3.75 per million output tokens, roughly half the standard rate.
  • Google has not released a frontier-level Gemini Pro model since early 2026, and the promised Gemini 3.5 Pro looks increasingly unlikely to appear.

Google has launched Gemini 3.8 Flash, its best reasoning and coding model yet according to the company, and it is already the third Flash release in just six weeks. That pace is worth pausing on.

What does Gemini 3.8 Flash actually do?

It handles two jobs. The standard Gemini 3.8 Flash is a general-purpose model built for everything from writing code to running agentic tasks, software that can carry out multi-step jobs on its own without a human clicking through each step. Think of it as the workhorse that sits behind developer tools and AI assistants.

The second version, Gemini 3.8 Flash Cyber, runs on the same underlying technology but has been fine-tuned for cybersecurity work: spotting vulnerabilities in software and suggesting fixes. It is aimed squarely at security teams and developers who need an AI that understands attack patterns, not just grammar.

What does this cost developers?

For now, cheap. Google is offering access through its API, the programming interface that lets developers connect their apps to the model, at an introductory rate through the end of 2025. That works out to $0.75 per million input tokens and $3.75 per million output tokens. A token is roughly three-quarters of a word, so a million tokens covers a very large amount of text.

The standard rate after the promotional period ends rises to $1.50 and $7.50 per million tokens respectively.

As Ars Technica noted, those prices probably reflect competitive pressure. Several rival AI labs have cut token costs recently to keep businesses from drifting away, and Google is matching that move.

Where is Gemini Pro?

Good question, and nobody at Google is answering it directly. The company has not shipped a frontier-level Gemini Pro model since early 2026. Frontier models are the biggest, most capable AI systems a lab produces; Pro sits above Flash in Google's lineup in terms of raw power. A Gemini 3.5 Pro was promised, but three Flash launches in six weeks suggests Google may have quietly shelved it in favour of iterating on the faster, cheaper Flash line instead.

For ordinary users who rely on Gemini inside Google products, Docs, Gmail, Search, the practical difference is small for now. Flash models are genuinely capable. But businesses that were waiting for a more powerful Pro upgrade before committing to Google's AI stack may want to reassess their timelines.

Common questions

Will this affect Gemini inside Google's apps?

Not immediately in a visible way. Flash models power many everyday Gemini features already, and 3.8 Flash is faster and smarter than its predecessors, so responses may feel snappier.

Is the half-price deal worth locking in now?

The introductory rate lasts through the end of 2025. Developers building tools that call the API heavily could save meaningfully, though Google may release yet another model before the pricing changes anyway.

© 2026 AI2Day