Gemini 3.8 Flash icon
Gemini 3.8 FlashGold Verified Icon
#9 in LLM models
4.4/5
« A fast and cost-effective model for code and long-running agent tasks, with clear improvements in multi-step reasoning. Available at the same price as the previous version through the end of 2026 »
Verified Icon
Verified Tool
Freemium 197895

Gemini 3.8 Flash: Google's budget workhorse punches at frontier weight

Gemini 3.8 Flash is Google's workhorse model, built for long-horizon software engineering, autonomous agents and multi-step reasoning in specialist fields. It shipped on September 2, 2026, three weeks after Gemini 3.7 Flash, and keeps the same launch rate of $0.75 per million input tokens. On DeepSWE v1.1 it now solves engineering problems end to end better than most larger frontier models. Paying Gemini app subscribers get it too, alongside AI Mode in Search and Gemini in Sheets.

Pros
  • Near-frontier results at a fraction of the cost
  • One-million-token window with audio and video input
  • Big jump in autonomous coding on DeepSWE v1.1
  • Adjustable effort levels to keep token spend in check
Cons
  • Higher token consumption on complex tasks by design
  • Cyber variant locked behind a restricted access program
  • Gemini app access requires a paid subscription

Gemini 3.8 Flash grinds through long tasks and checks its own work

Gemini 3.8 Flash takes smaller reasoning steps, calls tools iteratively and verifies its output along the way, a deliberate design choice for long-running jobs. The context window reaches one million tokens, with text, images, audio and video accepted as input and 65,536 tokens of output. You hand it a full repository, it reads, patches, runs the tests and loops until everything goes green.

That stubbornness shows up in the numbers. On DeepSWE v1.1, a long-horizon software engineering benchmark, it climbs to roughly 71% against 65.3% for version 3.7, and it posts 54.9% on HLE-Verified. Running autonomous AI agents is where it earns its keep, with Google also reporting gains on professional tests such as Vals Finance Agent V2 and Harvey's legal benchmark.

Three Flash releases in six weeks, plus a locked-down security twin

Google has been shipping a fresh Flash model roughly every three weeks since late July 2026, and Gemini 3.8 Flash is the third in that run. The pace is almost dizzying, especially with the Gemini 3.5 Pro announced at I/O in May 2026 still nowhere in sight.

Same day, second announcement. Gemini 3.8 Flash Cyber, restricted to the Fairwind Program, hunts security flaws on its own. The Chrome security team credits it with 2.6 times more correct vulnerability patches than much larger commercial models, while Wiz measured higher recall on its penetration testing benchmark at a sharply lower cost.

ModelReleaseWhat changed
Gemini 3.6 FlashJuly 21, 2026Start of the rapid release cycle
Gemini 3.7 FlashAugust 13, 2026Focus shifted to coding
Gemini 3.8 FlashSeptember 2, 2026Long-horizon agents, expert-level analysis
Gemini 3.8 Flash CyberSeptember 2, 2026Security work, gated behind Fairwind

Gemini 3.8 Flash in the app, in Sheets and through the API

Consumer access runs through the Gemini app, AI Mode in Google Search and Gemini in Google Sheets, with a Google AI Pro or Ultra subscription. Developers find it in the Gemini API, Google AI Studio, Android Studio, Antigravity and Gemini Enterprise.

Then comes the bill. The launch rate matches version 3.7 at $0.75 per million input tokens and $3.75 per million output tokens, thinking tokens included (that detail hides in a footnote of Google's own benchmark table). Both figures double on January 1, 2027, while batch mode cuts the tab in half. Before locking a budget, check Google's official rate card, since it shifts with every release.

Frequently asked questions

Is Gemini 3.8 Flash free?

No, not in the Gemini app, where selecting it requires a Google AI Pro or Ultra subscription. Developers pay per token through the API, at an introductory rate that runs until December 31, 2026. Google AI Studio remains the simplest place to try it before committing any budget.

Why does Gemini 3.8 Flash use more tokens than 3.7 Flash?

Because it works harder, in Google's own words. On complex tasks it runs extra reasoning steps and calls tools repeatedly, pushing cost per task up by roughly 40% even though the per-token price is unchanged. You can dial down the effort level, or simply keep running version 3.7.

Who can access Gemini 3.8 Flash Cyber?

Only trusted government agencies, critical infrastructure operators and software maintainers, through the new Fairwind Program. Google is keeping tight control of a model that discovers vulnerabilities autonomously, after its cloud research team found a critical flaw in under two hours, a job that usually takes months of work.

What is the knowledge cutoff of Gemini 3.8 Flash?

March 2026 at the latest, with cutoff dates ranging from January 2025 to March 2026 depending on the domain. For fresher information, the model leans on external tools such as web search, billed separately through the API.

Verdict: For running coding or research agents all day long without frontier-tier invoices, this Flash has become a very strong pick, and teams already living inside Google's ecosystem will get the most from it.

★ Featured AI Tools ★
AI Alternatives for Gemini 3.8 Flash
Free
Paid
Paid
Paid
Free
Paid
Freemium
Paid