<!-- mobian-agent-page publisher="dailydev" canonical="https://daily.dev/posts/gemini-3-8-flash-is-google-s-third-flash-drop-in-six-weeks-and-it-s-gunning-for-claude-wyexjkmxw" -->

---
title: Gemini 3.8 Flash is Google&#x27;s third Flash drop in six...
description: Google released Gemini 3.8 Flash, its third Flash model in six weeks, now generally available across Gemini, Google AI Studio, and the API. It scores 71% on...
canonical: https://daily.dev/posts/gemini-3-8-flash-is-google-s-third-flash-drop-in-six-weeks-and-it-s-gunning-for-claude-wyexjkmxw
twitter:card: summary_large_image
twitter:site: @dailydotdev
og:type: website
og:site_name: daily.dev
og:title: Gemini 3.8 Flash is Google&#x27;s third Flash drop in six weeks, and it&#x27;s gunning for Claude | daily.dev
og:description: Google released Gemini 3.8 Flash, its third Flash model in six weeks, now generally available across Gemini, Google AI Studio, and the API. It scores 71% on...
og:url: https://daily.dev/posts/gemini-3-8-flash-is-google-s-third-flash-drop-in-six-weeks-and-it-s-gunning-for-claude-wyexjkmxw
og:image: https://api.daily.dev/og/posts/WyexJKmxW.png
og:image:alt: Gemini 3.8 Flash is Google&#x27;s third Flash drop in six weeks, and it&#x27;s gunning for Claude
og:image:width: 1200
og:image:height: 630
og:locale: en
---

> ## Documentation Index
> Fetch the complete documentation index at: https://daily.dev/llms.txt
> Use this file to discover all available pages before exploring further.

# Gemini 3.8 Flash is Google's third Flash drop in six weeks, and it's gunning for Claude

**[Trends](https://daily.dev/sources/trends)** · 2 min read · 2 upvotes · 0 comments

## Summary

Google released Gemini 3.8 Flash, its third Flash model in six weeks, now generally available across Gemini, Google AI Studio, and the API. It scores 71% on DeepSWE 1.1 versus Claude Opus 5's 74%, at a much lower price: $0.75/1M input tokens and $3.75/1M output tokens (including thinking tokens) through end of 2026, doubling January 1, 2027. It's now the default model in Antigravity and Managed Agents, and was spotted in production before the official announcement. Google's Phil Schmid frames the rapid cadence as focused on agentic autonomy and long-horizon tasks, while Primeagen sarcastically urged people to check the model card before hype takes over.

## Content

Three Flash releases in six weeks. That's Google's current pace, and Gemini 3.8 Flash is the latest — generally available now across Google AI Studio, the Gemini API, and GCP's Agent Studio, priced at $0.75/$3.75 per million tokens through the end of 2026.

The benchmarks are legitimately interesting. On HLE-Verified, 3.8 Flash scores 54.9% against Claude Opus 5's 54.4%. On Harvey's Legal Agent Benchmark — a brutal all-or-nothing eval where every criterion has to pass — it scores 10.0% to Opus 5's 6.7%. On the Cursor bench, 69.9% at $2.38 per task. These aren't Flash-tier numbers. They're frontier-adjacent numbers at Flash-tier prices.

The model itself takes a different approach than its predecessors: smaller steps, more frequent self-verification. That means higher token usage on complex tasks, but Google's Phil Schmid argues those extra tokens are doing real work — running tests, checking outputs, catching mistakes before they compound. Early access users seem to agree. Omar Sanseviero called it "a very well rounded agentic model" after building with it for a few days. Ethan Mollick was more measured: "a very good Flash model, but not equivalent to a frontier model."

The discourse, though, has landed somewhere more uncomfortable for Google. Theo's take — "We're never getting a new Gemini Pro model huh" — is getting traction because it's hard to argue with. And Hesamation put it bluntly: "Google cooked so hard with 3.8 Flash that they can't ship a Pro that justifies the higher price."

That's the real story here. Google has been so aggressive on Flash that it's painted itself into a corner. What does a Pro model even offer at this point unless it's a genuinely different capability tier? The Flash line is eating the product roadmap from below.

Meanwhile, Vercel is running 50% off on AI Gateway access to the model through December 31st, and it's already the default in Managed Agents. The "Gemini 3.8 Flash Cyber" variant — focused on security patching — is also drawing attention, sitting on the Pareto frontier for vulnerability remediation according to early evals.

The model is good. The pricing is aggressive. The Pro question is getting louder.

## Questions this post answers

### What is the pricing for Gemini 3.8 Flash and when does it change?

Gemini 3.8 Flash costs $0.75 per 1M input tokens and $3.75 per 1M output tokens, including thinking tokens, through the end of 2026. That pricing doubles starting January 1, 2027, creating urgency for teams building agentic workloads to lock in the lower rate before then.

_Teams budgeting agentic workloads around token costs can track Gemini pricing shifts on daily.dev._

### How does Gemini 3.8 Flash compare to Claude Opus 5 on coding benchmarks?

Gemini 3.8 Flash scores 71% on DeepSWE 1.1 versus Claude Opus 5's 74%, a small gap given Flash's much lower price. Google positions the model as taking smaller steps and verifying work more often, generating higher token counts that are meant to catch errors before they compound in agentic tasks.

_Developers weighing Claude against Gemini for coding agents can follow these comparisons on daily.dev._

### Where is Gemini 3.8 Flash deployed as the default model?

Gemini 3.8 Flash is now the default model in Antigravity and Managed Agents, signaling Google is betting on it for production agentic pipelines rather than just benchmark showcases. It was also spotted in Agent Studio on GCP and Cloud Console quotas before the official announcement went live.

_Anyone standing up production agent pipelines can keep tabs on default model changes via daily.dev._

## Community take

How the wider developer community reacted, aggregated from 5 discussions and 69 comments across x (as of 2026-09-02).

**TL;DR:** Reactions to Gemini 3.8 Flash are mostly enthusiastic about its speed, agentic coding feel, and aggressive pricing, but a recurring thread questions whether Google's rapid Flash cadence makes the pricier Pro tier pointless, and some are skeptical benchmarks translate to real-world performance.

**Sentiment:** 40% positive · 45% mixed · 15% skeptical

**The case for**

- Coding agent workflows feel noticeably more natural, not just improved on benchmarks
- Context window handling on long multi-turn tool use has improved
- The Cyber variant's placement on the Pareto frontier for vulnerability patching stands out as a real shift
- Fast, cheap access via gateways lets developers try new Flash models without rewriting SDKs

**The pushback**

- Flash beating Pro-tier benchmarks undermines the case for paying for the Pro tier at all
- Benchmarks don't necessarily reflect real-world results
- Skepticism that Google models remain heavily censored
- One commenter joked it still can't handle basic tasks like reading emails, suggesting 'impressive' claims are overstated

**By community**

- x (mixed): Overall enthusiastic about speed, pricing, and agentic capability, tempered by pointed jokes and doubts about whether the Pro tier still has a reason to exist.

**Hottest debate:** Whether Flash's rapid improvement is making Google's higher-priced Pro tier obsolete.

**Open questions**

- Which specific domain or task type is Gemini 3.8 Flash still weakest at?
- How does the Cyber variant perform on real CVE patching in a live repository rather than benchmarks?
- Is the discounted gateway pricing actually cheaper than alternatives like OpenRouter?

**Highlights**

> @Hesamation Gemini 3.8 Flash is mogging so hard they've had to hide the Gemini 3.1 Pro stats from the table. How funny would it be if one day Flash model starts beating Fable class for real.
> — [MrMSpencer on x · 1 points](https://x.com/MrMSpencer/status/2095184619997958507)

> @Hesamation That ceiling problem is real, Flash beating opus 5 on HLE makes the Pro tier a harder pitch than any benchmark gap can fix.
> — [OussPoly on x](https://x.com/OussPoly/status/2095180316260327559)

> @osanseviero "Well rounded" is the giveaway a model shipped — nobody says that about the one that's actually great at one thing and mediocre at nine others. Curious which domain it's quietly worst at.
> — [ukrroot on x](https://x.com/ukrroot/status/2095192084130984203)

> @omarsar0 The cyber variant angle is interesting but the real story is deployment speed. If Gemini ships cyber-specific variants while others are still doing safety reviews, the window between capability and production narrows to days. Thats the actual competitive moat — not the benchmark.
> — [gunbit\_ on x](https://x.com/gunbit_/status/2095189834381787201)

> @Hesamation Benchmarks not equal to real-world results.
> — [synthshareai on x](https://x.com/synthshareai/status/2095198268850323836)

**Source threads**

- [x](https://x.com/osanseviero/status/2095180888921276643) · 0 points · 26 comments
- [x](https://x.com/Hesamation/status/2095175925507715209) · 1 points · 11 comments
- [x](https://x.com/rauchg/status/2095195988550168936) · 0 points · 26 comments
- [x](https://x.com/omarsar0/status/2095177610930098670) · 0 points · 6 comments
- [x](https://x.com/kloss_xyz/status/2095202250612486454) · 0 points · 0 comments

---

Tags: [#google](https://daily.dev/tags/google), [#llm](https://daily.dev/tags/llm), [#claude](https://daily.dev/tags/claude), [#google-gemini](https://daily.dev/tags/google-gemini), [#agentic-ai](https://daily.dev/tags/agentic-ai)

[View this post on daily.dev](https://daily.dev/posts/gemini-3-8-flash-is-google-s-third-flash-drop-in-six-weeks-and-it-s-gunning-for-claude-wyexjkmxw)

```json
{"@context":"https://schema.org","@graph":[{"@type":"Organization","@id":"https://daily.dev/#organization","name":"daily.dev","url":"https://daily.dev","logo":{"@type":"ImageObject","url":"https://daily.dev/apple-touch-icon.png","width":180,"height":180},"sameAs":["https://twitter.com/dailydotdev","https://github.com/dailydotdev","https://www.linkedin.com/company/daily-dev-ltd"]},{"@type":"WebSite","@id":"https://daily.dev/#website","url":"https://daily.dev","name":"daily.dev","publisher":{"@id":"https://daily.dev/#organization"},"potentialAction":{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https://daily.dev/search?q={search_term_string}"},"query-input":"required name=search_term_string"}}]}
{"@context":"https://schema.org","@type":"TechArticle","headline":"Gemini 3.8 Flash is Google's third Flash drop in six weeks, and it's gunning for Claude","url":"https://daily.dev/posts/gemini-3-8-flash-is-google-s-third-flash-drop-in-six-weeks-and-it-s-gunning-for-claude-wyexjkmxw","mainEntityOfPage":{"@type":"WebPage","@id":"https://daily.dev/posts/gemini-3-8-flash-is-google-s-third-flash-drop-in-six-weeks-and-it-s-gunning-for-claude-wyexjkmxw"},"datePublished":"2026-09-02T15:43:28.472Z","dateModified":"2026-09-02T18:34:48.622Z","description":"Google released Gemini 3.8 Flash, its third Flash model in six weeks, now generally available across Gemini, Google AI Studio, and the API. It scores 71% on...","image":"https://pbs.twimg.com/media/HROPESqbYAAl7IM.jpg","thumbnailUrl":"https://pbs.twimg.com/media/HROPESqbYAAl7IM.jpg","isAccessibleForFree":true,"articleSection":"Trends","inLanguage":"en","publisher":{"@type":"Organization","name":"daily.dev","url":"https://daily.dev","logo":{"@type":"ImageObject","url":"https://daily.dev/apple-touch-icon.png","width":180,"height":180}},"author":{"@type":"Organization","name":"Trends","logo":"https://media.daily.dev/image/upload/s--ZfSp3asX--/f_auto,q_auto/v1780996004/logos/trends?_a=BAMAMiWQ0","url":"https://daily.dev/sources/trends"},"commentCount":0,"discussionUrl":"https://daily.dev/posts/gemini-3-8-flash-is-google-s-third-flash-drop-in-six-weeks-and-it-s-gunning-for-claude-wyexjkmxw","interactionStatistic":[{"@type":"InteractionCounter","interactionType":{"@type":"LikeAction"},"userInteractionCount":2},{"@type":"InteractionCounter","interactionType":{"@type":"CommentAction"},"userInteractionCount":0}],"keywords":"google,llm,claude,google-gemini,agentic-ai","timeRequired":"PT2M"}
{"@context":"https://schema.org","@type":"BreadcrumbList","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https://daily.dev"},{"@type":"ListItem","position":2,"name":"Trends","item":"https://daily.dev/sources/trends"},{"@type":"ListItem","position":3,"name":"Gemini 3.8 Flash is Google's third Flash drop in six weeks, and it's gunning for Claude"}]}
{"@context":"https://schema.org","@type":"FAQPage","@id":"https://daily.dev/posts/gemini-3-8-flash-is-google-s-third-flash-drop-in-six-weeks-and-it-s-gunning-for-claude-wyexjkmxw#faq","mainEntity":[{"@type":"Question","name":"What is the pricing for Gemini 3.8 Flash and when does it change?","acceptedAnswer":{"@type":"Answer","text":"Gemini 3.8 Flash costs $0.75 per 1M input tokens and $3.75 per 1M output tokens, including thinking tokens, through the end of 2026. That pricing doubles starting January 1, 2027, creating urgency for teams building agentic workloads to lock in the lower rate before then. Teams budgeting agentic workloads around token costs can track Gemini pricing shifts on daily.dev."}},{"@type":"Question","name":"How does Gemini 3.8 Flash compare to Claude Opus 5 on coding benchmarks?","acceptedAnswer":{"@type":"Answer","text":"Gemini 3.8 Flash scores 71% on DeepSWE 1.1 versus Claude Opus 5's 74%, a small gap given Flash's much lower price. Google positions the model as taking smaller steps and verifying work more often, generating higher token counts that are meant to catch errors before they compound in agentic tasks. Developers weighing Claude against Gemini for coding agents can follow these comparisons on daily.dev."}},{"@type":"Question","name":"Where is Gemini 3.8 Flash deployed as the default model?","acceptedAnswer":{"@type":"Answer","text":"Gemini 3.8 Flash is now the default model in Antigravity and Managed Agents, signaling Google is betting on it for production agentic pipelines rather than just benchmark showcases. It was also spotted in Agent Studio on GCP and Cloud Console quotas before the official announcement went live. Anyone standing up production agent pipelines can keep tabs on default model changes via daily.dev."}}]}
```

