Install AgentBook a Call
All posts

August 14, 2026 · 3 min read

Google Releases Gemini 3.7 Flash: Better Agents and Half the Cost

AI NewsAutomationGemini

Google has released Gemini 3.7 Flash, positioning it as the company's most intelligent "workhorse" model for coding and agent-driven workflows. Arriving just three weeks after Gemini 3.6 Flash, the update introduces major improvements to complex reasoning, tool calling, and automated execution.

Critically for businesses looking to scale their AI operations, Google is offering 3.7 Flash at an introductory price through the end of the year that halves the original cost of 3.6 Flash. The new pricing sits at $0.75 per million input tokens and $3.75 per million output tokens.

Measurable Gains in Business and Logic Workflows

For business operations, the most relevant metrics from Google’s announcement center on document processing and task automation. On the GDP.pdf benchmark—which evaluates a model's ability to process complex, knowledge-dense documents in fields like finance and law—3.7 Flash scored 34.0%, up from 3.6 Flash's 22.0%.

The model also demonstrated a massive jump on AutomationBench, scoring 30.4% compared to its predecessor's 17.0%. Google attributes this to the model applying more effort into multi-step planning, adapting better to roadblocks, and executing tool calls with greater discipline. The company notes this translates directly to less manual oversight and fewer retries across workflows.

Advancements in Software and Web Development

Beyond administrative logic, 3.7 Flash shows strong gains in coding accuracy and issue resolution. It improved performance in generating production-ready code, moving from 34.4% to 43.6% on FrontierCode 1.1 Main, and jumping from 49.0% to 65.3% on DeepSWE v1.1.

In web development, the model generates functional layouts and feature-complete apps with fewer prompts. It achieved an Elo score of 1588 on Arena.ai’s WebDev Arena (up from 1538). Google highlighted several advanced use cases, including combining 3.7 Flash with Nano Banana to generate real-time 3D game assets, and orchestrating sub-agents via Gemini Omni to build interactive parallax landing pages in a single shot.

Smarter Tools for Google Workspace

Google is immediately integrating 3.7 Flash into Gemini Spark, its 24/7 personal AI agent available to Google AI Pro and Ultra subscribers in over 160 countries. The integration focuses heavily on knowledge work within Google Workspace apps. Because of the model's improved tool-use capabilities, it can execute multi-skill workflows like consolidating files, drafting emails, and updating status documents with higher accuracy.

What This Means for SMB Automation

When a major provider releases a frontier-class model that drops in price while nearly doubling its performance on automation benchmarks, it shifts the operational math for small and mid-sized businesses.

  • High-Volume Document Extraction: The improvements in dense document reasoning mean SMBs can more reliably automate data extraction from complex PDFs, such as annual reports, legal agreements, or non-standard vendor invoices. A model that hallucinates less on dense text requires fewer human reviewers in the loop.
  • Reliable Multi-Step Agents: An AI that can better handle roadblocks and clarify intent is a prerequisite for autonomous operations. When building automations that span multiple platforms—like moving customer data from an inbox, formatting it, and updating a CRM—the typical failure point is poor tool execution. The 3.7 Flash update specifically targets this execution gap, allowing businesses to trust AI with multi-stage workflows with less manual correction.
  • Lower Barrier to Scale: At $0.75 per million input tokens, processing high volumes of text in the background is highly cost-effective. Admin and operations tasks that were previously too expensive or error-prone to automate with earlier models are now viable for smaller budgets.

Inspired by this source.

Stop doing the busywork yourself

A dedicated AI technician — backed by a full engineering team — automates your admin inside the tools you already use.