All posts

Google ships Gemini 3.7 Flash at half-price for agents

Huma ShaziaAugust 14, 2026 at 5:16 AM4 min read
Google ships Gemini 3.7 Flash at half-price for agents

Google released Gemini 3.7 Flash today, a model built for software coding and autonomous business workflows, but gave no timeline for its delayed flagship Gemini 3.5 Pro. The company is pricing the new model at half the cost of its predecessor through year-end: 75 cents per million input tokens and $3.75 per million output tokens.

Google Gemini AI model branding graphic showing the company's artificial intelligence initiative
Image for Google unveils Gemini 3.7 Flash AI model for coding, agent workflows

The launch comes three weeks after Gemini 3.6 Flash and positions 3.7 as a lower-cost option for companies building AI systems that plan tasks, use software tools, and complete multi-step workflows without constant human oversight. Google claims improved performance on debugging, issue resolution, and production-ready code generation, though it did not publish benchmark comparisons.

Advertisements

What does Gemini 3.7 Flash actually do?

Google Shipped Gemini 3.7 Flash And OpenAI Made GPT-5.6 14x Faster!

Google is targeting agent workflows, the class of AI applications where a model doesn't just respond to prompts but takes actions. Think automated code review that opens pull requests, customer service bots that process refunds, or data pipelines that debug themselves.

The model is available immediately through Gemini Spark, Google's subscription AI agent service bundled with Google AI Pro and Ultra tiers in over 160 countries. Enterprise customers can access it via Google Cloud's Vertex AI.

50%
Price reduction vs. Gemini 3.6 Flash through end of 2026

The Pro model remains missing

Investors and developers have been waiting on Gemini 3.5 Pro since Google said in July it was testing with partners and arriving "soon." That was six weeks ago. The company offered no update today.

The delay matters because Pro models are Google's answer to Anthropic's Claude and OpenAI's GPT-4 class offerings. Flash models trade capability for speed and cost. Shipping Flash after Flash without Pro suggests DeepMind is struggling to close the capability gap at the high end, or is prioritizing agent-ready models over raw reasoning power.

DeepMind's turbulent month

Google announced a leadership overhaul at DeepMind last week. CEO Demis Hassabis stepped aside for his deputy, Koray Kavukcuoglu. Simultaneously, the two original technical co-leads of Gemini left to start their own company.

Co-founder Sergey Brin has reportedly urged key AI staff to go "all in" on Gemini as Alphabet tries to close ground with OpenAI and Anthropic. CEO Sundar Pichai defended Google's AI strategy on the July earnings call, pushing back on concerns that repeated delays signal deeper problems.

Also Read
Why AI agents are making specialists obsolete

Explains the agent workflow trend Gemini 3.7 Flash targets

Who should care about the pricing

The introductory pricing is aggressive. At 75 cents per million input tokens, Gemini 3.7 Flash undercuts many competitors for high-volume agent applications. The catch: this rate expires at year-end, and Google hasn't said what happens after. Companies building on Flash need to budget for a potential 2x price jump in January.

For teams evaluating automation workflows, tools like Zapier, Make, and n8n already integrate with Google's AI APIs. The lower per-token cost changes the math on which tasks are worth automating.

ℹ️

Disclosure

Some links in this post are affiliate links — Logicity earns a commission if you sign up, at no extra cost to you. We only link products we have used or actively recommend.

ℹ️

Logicity's Take

Shipping a minor Flash revision while Pro stays frozen is not a confidence signal. Google is winning on distribution (Gemini Spark in 160+ countries) and price, but the model that proves DeepMind can compete at the frontier keeps slipping. For CTOs picking an AI vendor: Flash is fine for cost-sensitive agent tasks, but don't bet your core product on Google's roadmap until Pro ships and benchmarks.

Also Read
1% AI error rate means 100 daily problems at scale

Relevant to anyone deploying AI agents in production workflows

ℹ️

Need Help Implementing This?

Evaluating AI coding tools or building agent workflows? Reach out to the Logicity team for vendor-neutral guidance on model selection and integration strategy.

Source: Tech-Economic Times / ET

H

Huma Shazia

Senior AI & Tech Writer

Produced with AI assistance and reviewed by the Logicity editorial team. Learn more in our Editorial Policy.

Related Articles