Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more

Google launched Gemini 3.8 Flash, arriving just a few weeks after its predecessor. The company claims the new model "works harder" than Gemini 3.7 Flash by performing more reasoning steps on complex tasks and "calling tools iteratively." It has the same introductory pricing as 3.7 Flash, $0.75 per million input tokens and $3.75 per million…
Opening of the original on The Verge AI
Summary
Google released Gemini 3.8 Flash, a new AI model that performs more reasoning steps and calls tools iteratively for complex tasks. While its introductory pricing matches its predecessor, Google warns that 3.8 Flash may use more tokens to maximize performance, potentially increasing costs for users. Developers can opt to continue using Gemini 3.7 Flash to control token usage.
Why it matters
Google's Gemini 3.8 Flash offers enhanced reasoning and tool usage, aiming for higher performance on complex tasks. This update directly impacts developers integrating Gemini into their applications, presenting a trade-off between potential cost increases and improved capabilities. Competitors like OpenAI and Anthropic also focus on model efficiency and performance. Developers must decide if the enhanced performance justifies the potential for higher token consumption. Future updates will likely focus on balancing this performance-cost equation.
Related: OpenAI: GPT‑6 Astra · Anthropic: GPT‑6 Astra · Microsoft: Meet MAI-Transcribe-2: A faster and more accurate speech recognition model · Meta: Meta Muse Code & Muse Spark Course – Build AI Agents, APIs, and Full-Stack Apps · xAI: Sam Altman "AGI by December"
Rated low: routine. Worth knowing, not worth rearranging your day for.
- llm-gemini 0.34Simon Willison · Web · Sep 2, 2026
- Google releases Gemini 3.8 Flash, its third Flash model in six weeksArs Technica AI · Web · Sep 2, 2026
- The Most Overhyped and Underhyped New AI ModelsMatt Wolfe · Web · Sep 2, 2026
- The Most Overhyped and Underhyped New AI ModelsMatt Wolfe · Web · Sep 2, 2026
- The Most Overhyped and Underhyped New AI ModelsMatt Wolfe · Web · Sep 2, 2026
- GOOGLE IS BACK! (Gemini 3.8 Flash)Matthew Berman · Web · Sep 3, 2026
- llm-gemini 0.34Simon Willison
- Introducing Gemini 3.8 Flash and 3.8 Flash CyberGoogle DeepMind Blog
- Introducing agentic video understanding with GeminiGoogle DeepMind Blog
- Google releases Gemini 3.8 Flash, its third Flash model in six weeksArs Technica AI
- Gemini 3.8 Flash rolling out three weeks after last release9to5Google
- Introducing Gemini 3.8 Flash and 3.8 Flash CyberThe Keyword: Gemini
Questions people ask
- What is new in Gemini 3.8 Flash?
- Gemini 3.8 Flash performs more reasoning steps and calls tools iteratively on complex tasks compared to its predecessor.
- Will Gemini 3.8 Flash cost more?
- While the introductory price is the same as Gemini 3.7 Flash, Google warns that the model might use more tokens to maximize performance, potentially increasing costs.
- Can I still use Gemini 3.7 Flash?
- Yes, developers can continue using Gemini 3.7 Flash if they want to minimize token usage.
More from The Verge AI 1 more
Page generated Sep 3, 2026. Summaries are Subvolts' own; the story belongs to The Verge AI.







