All AI News
    The Verge AIWednesday, September 2, 2026 2 min read
    AI

    Google says its new Gemini 3.8 Flash model 'works harder' but might cost more

    Gemini 3.8 Flash's 'work harder' design means same list price but potentially higher real-world bills.

    Key takeaways
    • 01Google's rapid cadence continues with Gemini 3.8 Flash, weeks after 3.7 Flash.
    • 02The upgrade adds deeper reasoning chains and iterative tool calls on complex tasks—but the efficiency trade-off is explicit: Google acknowledges the model may consume significantly more tokens at higher effort settings.
    • 03Nominal pricing holds at $0.75/$3.75 per million input/output tokens, yet actual costs could climb.
    • 04Google is preserving 3.7 Flash access specifically for teams prioritizing token efficiency over raw capability.
    Koko brief

    Gemini 3.8 Flash's 'work harder' design means same list price but potentially higher real-world bills.

    Google's rapid cadence continues with Gemini 3.8 Flash, weeks after 3.7 Flash. The upgrade adds deeper reasoning chains and iterative tool calls on complex tasks—but the efficiency trade-off is explicit: Google acknowledges the model may consume significantly more tokens at higher effort settings. Nominal pricing holds at $0.75/$3.75 per million input/output tokens, yet actual costs could climb. Google is preserving 3.7 Flash access specifically for teams prioritizing token efficiency over raw capability.

    Watch: Monitor token consumption closely in production before migrating workloads from 3.7 Flash—benchmark real cost delta, not just listed rates.

    In brief · from theverge.com

    Google launched Gemini 3.8 Flash , arriving just a few weeks after its predecessor . The company claims the new model "works harder" than Gemini 3.7 Flash by performing more reasoning steps on complex tasks and "calling tools iteratively." It has the same introductory pricing as 3.7 Flash, $0.75 per million input tokens and $3.75 per million output tokens, but could still end up costing users more. Google warns that "the model might use more tokens to maximize performance, especially at higher effort levels."

    Read the full article at theverge.com

    Google launched Gemini 3.8 Flash , arriving just a few weeks after its predecessor . The company claims the new model "works harder" than Gemini 3.7 Flash by performing more reasoning steps on complex tasks and "calling tools iteratively." It has the same introductory pricing as 3.7 Flash, $0.75 per million input tokens and $3.75 per million output tokens, but could still end up costing users more. Google warns that "the model might use more tokens to maximize performance, especially at higher effort levels." Developers can keep using Gemini 3.7 Flash if they want to minimize token usage. Gemini 3.8 Flash's launch was followed by some early … Read the full story at The Verge.

    Don't miss tomorrow's

    The Daily Pulse in your inbox each morning — sourced and linked.

    How often
    Keep going — across the app