Google Launches Gemini 3.7 Flash: Its Most Intelligent Workhorse Model Yet for Coding and Agents

2026-08-14·10 min read

On August 13, 2026, Google officially launched Gemini 3.7 Flash. According to the announcement posted by Tulsee Doshi, Senior Director of Product Management, the new model is described by the company as 'our most intelligent workhorse model yet for coding and agents.' Gemini 3.7 Flash delivers a significant performance improvement over the previous-generation 3.6 Flash while maintaining low cost, especially demonstrating higher accuracy and output quality in complex, multi-skill workflows. At the same time, Google also announced that its new personal AI agent Gemini Spark will use 3.7 Flash as its underlying driving model, providing 24/7 personal assistant services to Google AI Pro and Ultra subscribers across more than 160 countries and regions.

From a product positioning perspective, Gemini 3.7 Flash continues the Gemini Flash series' consistent 'high value workhorse model' route, but this time Google places particular emphasis on the two keywords 'agents' and 'coding.' This means the new model's optimization focus is no longer limited to traditional text generation or Q&A, but centers on the model's ability to execute complex tasks as an autonomous agent — including writing and debugging code, calling external tools, planning multi-step workflows, and maintaining task consistency and coherence over long-running sessions. For developers, this is a foundational model built for the 'agent era': it must not only give the right answer in a single conversation, but also be able to complete end-to-end tasks in real working environments, replacing human labor. Google specifically noted that Gemini 3.7 Flash shows significantly better accuracy and output quality than 3.6 Flash when handling tasks that require combining multiple skills, such as scenarios involving code understanding, document retrieval, and data processing simultaneously.

Alongside 3.7 Flash, Google also unveiled Gemini Spark, its new personal AI agent. According to the official introduction, Gemini Spark is a personal AI agent that runs 24/7, designed specifically for Google AI Pro and Ultra subscribers, covering more than 160 countries and regions. This means users can rely on Spark as an always-online personal assistant, letting it continuously handle various tasks in the background: managing schedules, organizing emails, booking services, tracking information, and even proactively discovering and solving potential problems when the user does not initiate a conversation. With 3.7 Flash as its underlying engine, Gemini Spark's reasoning ability, tool-use capability, and long-task execution ability are all guaranteed. Notably, Google emphasized Spark's positioning of 'working for the user' in its launch — marking Google's official entry into the 'personal AI agent' competition pushed by OpenAI, Anthropic and others, attempting to upgrade AI from a passively responding tool into an autonomous system that proactively completes tasks for users.

In terms of tool use, a key upgrade of Gemini 3.7 Flash is the significantly improved support for Google Workspace applications. Google Workspace is one of the most widely used productivity suites in the world, covering Gmail, Google Docs, Google Sheets, Google Slides, Google Calendar and other applications. The improvement in 3.7 Flash's Workspace tool use means the model can operate these applications more accurately and reliably through natural language instructions — for example, automatically composing and sending emails, batch processing data in spreadsheets, generating presentations, and coordinating meeting times. This capability is especially important for Gemini Spark: as a 24/7 personal agent, Spark needs to interact frequently with the user's work tools, and the accuracy and stability of tool use directly determine whether it can truly take on the responsibility of 'working for the user.' Google emphasized that this improvement makes the model more robust when handling complex workflows involving multiple Workspace applications, reducing error-prone calls and context loss.

In terms of pricing, Google has set an attractive introductory price for Gemini 3.7 Flash. The introductory pricing is valid until December 31, 2026; from January 1, 2027, input tokens are priced at $1.50 per million and output tokens at $7.50 per million. This pricing structure continues the Gemini Flash series' consistent 'low input cost, relatively higher output cost' strategy, and also means that during the introductory period users can experience the new model at a more favorable price ahead of time. Looking at Google's overall pricing strategy, the Flash series has always been positioned as the preferred model for 'cost-sensitive scenarios' — suitable for large-scale invocation, batch task processing, and enterprise applications that are sensitive to per-call costs. While maintaining this cost-performance advantage, 3.7 Flash raises performance to a level close to flagship models, making it extremely competitive on the key metric of 'performance/price ratio.' For developers building AI agent applications, a model that is both powerful and cheap means lower operating costs and a more viable business model.

🤔 Frequently Asked Questions

Q1: What is Gemini 3.7 Flash? What improvements does it have over 3.6 Flash?

Gemini 3.7 Flash is the new-generation Flash series model released by Google on August 13, 2026, described by the company as 'our most intelligent workhorse model yet for coding and agents.' Compared with the previous-generation 3.6 Flash, its core improvements are reflected in three aspects: first, it achieves significant performance gains while maintaining low cost; second, it demonstrates higher accuracy and output quality in complex, multi-skill workflows; third, its tool-use capability for Google Workspace applications is markedly improved. These improvements make 3.7 Flash especially suitable as the underlying driving model for AI agents, executing complex tasks such as coding, tool calling, and multi-step planning.

Q2: What is Gemini Spark? What is its relationship with 3.7 Flash?

Gemini Spark is Google's personal AI agent, positioned as a 24/7 personal assistant for Google AI Pro and Ultra subscribers across more than 160 countries and regions. It can continuously handle various tasks in the background such as schedule management, email organization, service booking, and information tracking, and even proactively discover and solve problems for users. Gemini 3.7 Flash is exactly the underlying driving model of Gemini Spark — Spark's reasoning, tool use, and long-task execution capabilities are all powered by 3.7 Flash. In short, 3.7 Flash is the 'engine,' while Gemini Spark is the 'whole vehicle' equipped with this engine.

Q3: How much does Gemini 3.7 Flash cost? When does charging begin?

Google has set introductory pricing for Gemini 3.7 Flash, valid until December 31, 2026, during which users can enjoy more favorable prices. From January 1, 2027, official pricing takes effect: input tokens at $1.50 per million and output tokens at $7.50 per million. This pricing structure continues the Flash series' characteristic of 'low input cost, relatively higher output cost,' making it highly competitive in cost-sensitive scenarios. For developers planning to deploy AI agents at scale in 2027, it is recommended to fully test the model's capabilities and plan cost budgets during the introductory period.

Q4: What impact does Gemini 3.7 Flash have on developers and the AI agent ecosystem?

The launch of Gemini 3.7 Flash has multiple impacts on developers and the AI agent ecosystem. First, it provides a model choice that combines high performance with low cost, lowering the barrier to building AI agent applications and allowing small and medium teams to deploy autonomous agents at affordable costs. Second, its focused optimization on coding and tool use directly responds to the current industry hotspot of 'AI agents' — developers can use a more reliable underlying model to build agents that truly complete tasks. Third, the improved Google Workspace tool use will drive more agent applications targeting office scenarios. Finally, with the launch of Gemini Spark, Google has joined the personal AI agent competition, which will accelerate the entire industry's transition from 'conversational AI' to 'autonomous agents.'

🛠️ Recommended Tools

  • JSON to CSV Converter - Analyze Gemini 3.7 Flash API pricing and performance comparison data, convert JSON-format model benchmark results to CSV for cost estimation
  • Percentage Calculator - Calculate the performance improvement of 3.7 Flash over 3.6 Flash, token cost changes, and deployment ROI
  • Word Counter - Count token estimates and word counts of model output text, optimize agent prompt design and content generation costs

Summary

The launch of Gemini 3.7 Flash marks another important move in Google's AI model competition. As the new-generation Flash model officially described as 'our most intelligent workhorse model yet for coding and agents,' 3.7 Flash achieves significant performance gains over 3.6 Flash while maintaining its low-cost advantage, and demonstrates higher accuracy and output quality in complex, multi-skill workflows. It is not only the underlying driver of Gemini Spark, the 24/7 personal AI agent, but also paves the way for AI agents to land in real office scenarios through improved tool use for Google Workspace applications. On pricing, Google uses introductory pricing to attract early users, with input tokens at $1.50 per million and output tokens at $7.50 per million from 2027. For developers and enterprises, Gemini 3.7 Flash provides a foundation for building agents that combines performance with cost-effectiveness, and as the personal AI agent competition intensifies, the entire industry is accelerating its transition from conversational AI to the era of autonomous agents.