PromptsRush
Prompts

Browse

All PromptsThe full curated libraryPrompts GalleryVisual, Pinterest-style browsingImage PromptsMidjourney, DALL·E & SDXLVideo PromptsRunway, Kling & SoraText & TemplatesChatGPT & Claude system prompts

Discover

CategoriesExplore prompts by topicAI ModelsBest prompts per modelPrompt PacksCommunity, passcode-protectedSubmit a PromptShare with the community

For Creators

Turn prompts into followers

Share passcode-protected prompt packs and grow your audience with Auto DM.

Start sharing
Marketplace

Explore

Shared PromptsPasscode-protected prompt packsAI SkillsNewInstallable Agent SkillsDesign SystemsNewLive themes & design tokens

Contribute

Submit a PromptPublish a prompt packSubmit a SkillShip an Agent SkillSubmit a DesignShare a design system

New · Skills

Teach your AI new tricks

Install ready-made skills for Claude, ChatGPT, Gemini, n8n & more.

Browse skills
Learn

Learning Tracks

Prompt EngineeringWrite prompts that deliverAI SkillsBuild & ship Agent SkillsAI AutomationWorkflows, agents & MCPDesign SystemsOn-brand UI with AI

More

Learning HubAll tracks · 40+ lessonsBlogGuides, news & deep diveseBooksPremium prompt packs & guides

100% Free

Learn AI, the practical way

From fundamentals to advanced across four hands-on tracks — no fluff.

Explore the hub
Blog
LoginSign Up
PromptsRush

The ultimate directory for finding, sharing, and managing production-ready AI prompts, system instructions, and advanced templates.

TwitterGitHubYouTubeInstagramEmail

Platform

  • Home
  • Browse Prompts
  • Marketplace
  • Skills
  • Categories
  • Submit a Skill

Top Categories

  • Image PromptPopular
  • Video Prompts
  • Text Templates

Company

  • Privacy Policy
  • Terms of Service
  • Contact Us

Subscribe on YouTube

New AI prompt & skills tutorials every week.

Subscribe

© 2026 PromptsRush. Crafted with & Passion.

All systems operational
HomeBlogNews
News

Google Gemini 3.8 Flash: What's New?

Gemini 3.8 Flash shipped 2 September — Google’s third Flash release in six weeks. Benchmark gains are modest, but the introductory price doubles on 1 January 2027, and a second model ships that you probably cannot use.

P
PromptsRushSeptember 2, 2026
•7 min read75 views

Advertisement

Google Gemini 3.8 Flash: What's New?

Google shipped Gemini 3.8 Flash on 2 September 2026 — its third Flash release in six weeks. The generation before it, 3.6 Flash, was still losing to Claude Opus 5 on most head-to-heads. It is faster to say what did not change: the price, the speed and the model class. Everything else moved.

Two things in this release matter more than the benchmark bumps. One is a scheduled price doubling that most coverage has skipped past. The other is a second model shipped alongside it, Gemini 3.8 Flash Cyber, which you almost certainly cannot use.

Everything below comes from Google's own announcement, model documentation and DeepMind benchmark pages, captured while writing.

What Actually Changed

Google frames 3.8 Flash as improving on 3.7 across software engineering, agentic tasks and multi-step reasoning in specialised domains. The behavioural description is the most useful part of the announcement:

The models "exhibit greater diligence — executing extra reasoning steps, and calling tools iteratively," especially on complex tasks.

That is a meaningful characterisation rather than marketing. A model that runs more reasoning steps and loops tools more persistently behaves differently in an agent harness: it finishes more tasks unattended, and it costs more tokens per task. If you are running 3.7 Flash on a fixed token budget, expect 3.8 to spend more of it per call — the same effect that makes agentic token spend hard to forecast on any model.

Benchmarks against 3.7 Flash

Benchmark3.8 Flash3.7 FlashGain
HLE-Verified (expert reasoning)54.9%53.6%+1.3pp
Vals Finance Agent v261.4%59.0%+2.4pp
Harvey Legal Agent10.0%8.8%+1.2pp

Google also cites DeepSWE v1.1, where it says 3.8 Flash "outperforms most larger frontier models in autonomously solving complex engineering problems end to end" — though the published chart does not carry exact figures.

Read those gains honestly: they are real but incremental, in the 1–2.5 point range. This is a refinement release, not a generational jump. What makes it interesting is that the gains arrive at the same price and speed as 3.7 — Google is holding cost flat while pushing quality, which over three releases in six weeks compounds into something significant.

The Price Doubles on 1 January 2027

A small cost and an enormous cost producing an identical result

This is the detail to plan around. Gemini 3.8 Flash launched at:

PeriodInput / 1M tokensOutput / 1M tokens
Now → 31 December 2026$0.75$3.75
From 1 January 2027$1.50$7.50

That is a 100% increase, already scheduled and published. The introductory rate matches 3.7 Flash, so anyone budgeting from today's invoice is modelling roughly four months of pricing that expires.

If you are running Gemini Flash at volume, build your 2027 forecast on $1.50 / $7.50 and treat the next four months as a discount rather than a baseline. At the post-January rate Flash is still cheap — but it is not the same trade against Claude Opus 5 and GPT-5.6, which is a comparison worth re-running once the discount lapses.

Specs and Capabilities

PropertyValue
Model IDgemini-3.8-flash
Input context1,048,576 tokens (~1M)
Max output65,536 tokens (64K)
Input modalitiesText, image, video, audio, PDF
OutputText only
Thinking levelslow, medium, high

Supported: caching, code execution, file search, function calling, structured outputs, search grounding, Google Maps grounding, URL context, batch and Flex inference, and computer use in preview.

Not supported: image generation, audio generation, and the Live API. Flash is a reasoning and tool-use model — for generative media you still reach for a different model.

One change that will break existing code

Thinking supports low, medium and high — but not minimal. If you have code passing a minimal thinking level to an earlier Flash model and you swap the model ID to 3.8, that call will not behave as before. It is a small change with a sharp edge, and it is the first thing to check when migrating.

Gemini 3.8 Flash Cyber

Shipped alongside 3.8 Flash is a cybersecurity-specialised variant, and it is the more unusual release of the two.

Google reports it surpasses 3.5 Flash Cyber and "significantly larger frontier models" on CyberGym, and that internal testing found real-world vulnerability discovery exceeding 70% across 20 programming languages.

Worth noting what Google reports honestly: on CWE-Bench patching, 3.8 Flash Cyber scores 47.2% against a leading frontier model's 47.8% — a narrow loss, published rather than omitted. That is a more credible presentation than a page of clean sweeps, and it tells you patching remains genuinely hard for every model.

You probably cannot access it. Flash Cyber is gated behind the new Fairwind Program, a restricted-access scheme for "trusted defenders" — government and critical infrastructure — by application. This is offence-defence asymmetry management: a model good at finding vulnerabilities is equally good at finding them for the wrong people. Treat announcements about it as industry news, not a product you can adopt this quarter.

Where You Can Use It

SurfaceAccess
Gemini APIGenerally available
Google AI StudioGenerally available
Android Studio, StitchGenerally available
Gemini appGoogle AI Pro and Ultra subscribers
AI Mode in Google SearchPro and Ultra subscribers
Gemini in Google SheetsPro and Ultra subscribers
Gemini EnterpriseAvailable
3.8 Flash CyberFairwind Program only, by application

Three Releases in Six Weeks

Three columns of near-identical height with a laser line across their tops

3.6 → 3.7 → 3.8 in six weeks is the story underneath the story, and it cuts both ways.

In your favour: quality improves without a price rise, and you inherit the gains by changing a model string. Against you: an eval suite validated against 3.7 is stale within weeks, prompts tuned to one model's diligence behave differently on the next — a prompt library built for 3.6 is not automatically valid on 3.8, and "latest" is a moving target in production.

The practical response is to pin explicit model IDs rather than aliases, keep a small regression suite you can rerun in an afternoon, and schedule model upgrades deliberately instead of drifting onto whatever is newest. A model that runs extra reasoning steps and loops tools more is exactly the kind of change that passes a smoke test and shifts your token bill.

Should You Upgrade from 3.7?

If you…Verdict
Run agentic workflows or tool loopsYes — diligence gains land hardest here
Do software engineering tasksYes — the headline improvement area
Work in finance or legal analysisYes — the two benchmarks Google leads with
Run simple classification or extractionOptional — 1–2pp gains will not show
Are on a tight per-task token budgetTest first — extra reasoning steps cost tokens
Pass a minimal thinking levelFix that call before switching
Are budgeting 2027 spendModel $1.50 / $7.50, and re-check it against the alternatives

The Verdict

Gemini 3.8 Flash is a refinement release that is worth taking, with one asterisk. The benchmark gains are modest — 1 to 2.5 points — but they arrive at unchanged price and speed, and the behavioural shift toward more diligent tool use is the kind of change that matters more in an agent loop than a leaderboard suggests.

The asterisk is the pricing. Introductory rates through 31 December 2026, doubling on 1 January 2027, is a substantial change to schedule into any forecast that outlives this year. It does not make Flash expensive — it makes it a different value proposition against the alternatives, which is exactly the comparison worth running before you commit.

Recommended · Genspark

Try Genspark — the AI super-agent

Genspark researches, plans and acts across the web for you — multi-step agentic workflows in one prompt.

Try Genspark Free

Affiliate link · We may earn a commission

Keep Reading

Claude Opus 5 vs Gemini 3.6 Flash covers how the previous generation stacked up, and Claude pricing explained breaks down the subscription side of the frontier tier. Or browse all guides and prompts on PromptsRush.

❓

Frequently Asked Questions

10 questions answered

2 September 2026 — Google's third Flash release in six weeks, following 3.6 and 3.7. It shipped alongside a second model, Gemini 3.8 Flash Cyber, which is restricted to the new Fairwind Program.
$0.75 per million input tokens and $3.75 per million output tokens — but that is introductory pricing that expires on 31 December 2026. From 1 January 2027 it doubles to $1.50 / $7.50. Budget your 2027 forecast on the higher figure.
Improvements in software engineering, agentic tasks and multi-step reasoning, at unchanged price and speed. Benchmarks: HLE-Verified 54.9% (from 53.6%), Vals Finance Agent v2 61.4% (from 59.0%), Harvey Legal 10.0% (from 8.8%). Google also describes greater diligence — extra reasoning steps and more iterative tool calling.
1,048,576 input tokens (roughly 1M) with a 65,536-token maximum output. It accepts text, image, video, audio and PDF input, but produces text output only — image and audio generation and the Live API are not supported.
A cybersecurity-specialised variant for vulnerability detection and automated patching, reporting over 70% real-world vulnerability discovery across 20 programming languages. Most people cannot access it — it is gated behind the Fairwind Program, a restricted scheme for trusted defenders in government and critical infrastructure, by application only.
One change to check: thinking supports low, medium and high, but minimal is not supported. If you pass a minimal thinking level to an earlier Flash model and simply swap the model ID, that call will not behave as before. Otherwise the API surface is consistent.
Likely yes, on complex work. Google describes the model as executing extra reasoning steps and calling tools iteratively — behaviour that finishes more tasks unattended but spends more tokens doing it. If you run on a fixed per-task budget, test before switching rather than assuming the flat price means flat cost.
Via the Gemini API, Google AI Studio, Android Studio, Stitch and Gemini Enterprise. On consumer surfaces it is available to Google AI Pro and Ultra subscribers in the Gemini app, AI Mode in Google Search, and Gemini in Google Sheets.
On Google's own published benchmarks it edges Claude Opus 5 and GPT-5.6 Sol on HLE-Verified (54.9% vs 54.4% and 54.5%) — but a 0.5-point spread is a statistical tie, not a win, and these are vendor-selected benchmarks. The genuine story is price: comparable scores at roughly a sixth of the frontier per-token cost.
Yes for agentic workflows, tool loops and software engineering, where the diligence improvements land hardest. It is optional for simple classification or extraction, where 1–2 point gains will not be visible. Either way, pin explicit model IDs rather than aliases — three releases in six weeks makes latest a moving target.
Back to Blog

Table of Contents

In this article

  • 1What Actually Changed
  • Benchmarks against 3.7 Flash
  • 2The Price Doubles on 1 January 2027
  • 3Specs and Capabilities
  • One change that will break existing code
  • 4Gemini 3.8 Flash Cyber
  • 5Where You Can Use It
  • 6Three Releases in Six Weeks
  • 7Should You Upgrade from 3.7?
  • 8The Verdict
  • 9Keep Reading

Recent Posts

22 Prompts to Improve Landing Page Conversion Rates

Sep 7 · 15 min

20 Prompts to Improve an Ugly AI-Generated Website

Sep 7 · 15 min

20 Lead Generation Tools with Top-Notch AI Integrations

Sep 7 · 21 min

25+ Vibe Coding Prompts and AI Tools for Building Beautiful Sites

Sep 5 · 16 min

How to Create a Landing Page using AI with Prompts

Sep 5 · 12 min

Category

News

Advertisement

You May Also Like

News

Google Gemini 3.8 Flash vs Opus 5 vs GPT 5.6

Sep 28 min
News

ChatGPT Statistics (2026): Usage, Trend, Market & Growth

Aug 257 min
C
News

Claude AI Stats 2026: User Growth, Market, Trends & Finance

Aug 138 min