News

Google I/O 2026: From Search Box to Code Editor, Google Goes Fully Agentic

2026-05-20 #Google#Google IO#Agents

Dozens of product updates were dropped all at once over the course of two hours. Google I/O 2026 once again took the tech world by storm with its "saturation release." Every major update points to a single word—Agent.

"The most cutting-edge Agents might have only reached 0.1% of the world's population. The real question is how to deliver these frontier capabilities at scale to a broader user base."
— Sundar Pichai

Google is taking on this challenge itself. Here is the answer it has delivered.


At a Glance: Seven Major Highlights of this I/O

No. Release Key Highlights
01 Gemini 3.5 Series Unveiled Flash released first, Pro going live next month
02 Gemini Spark Competes with OpenClaw, a 24/7 online personal Agent
03 Antigravity 2.0 Upgraded from a coding environment to a multi-Agent orchestration platform, directly competing with Claude Code and Codex
04 Gemini Omni Any modality in, any modality out, integrating Veo and Nano Banana
05 Gemini App Goes Fully Agentic New design language Neural Expressive, featuring built-in Daily Brief
06 Search and Workspace Products Go "Mini-Gemini" Search box receives its biggest upgrade in 25 years; Information Agents and Generative UI debut
07 Subscription Pricing Adjustments Added a new $100/month tier, reduced the top-tier Ultra from $250 to $200, and kept Pro at $19.99

1. 3.5 Flash: Bringing the "Agent Economics" to the Table

As the first model in the Gemini 3.5 series, 3.5 Flash outperforms Google's own previous-generation Gemini 3.1 Pro across almost all benchmarks. The improvement is especially pronounced in GDPval (a metric measuring "tasks with real economic value"), which Pichai specifically highlighted.

But what truly put it in the spotlight of the keynote was another set of metrics:

  • Speed: Output speed is 4x faster than other frontier models; the optimized version built into Antigravity 2.0 reaches up to 12x faster, with no loss in quality.
  • Price: Less than half the price of comparable frontier models.

Pichai translated this price-performance ratio directly into a bill for CIOs to see:

"If a leading enterprise processes 1 trillion tokens a day, migrating 80% of that workload from other frontier models to 3.5 Flash could save over $1 billion a year."

Google itself is the biggest beneficiary of this approach. In March of this year, Google was processing 500 billion tokens internally per day; now it exceeds 3 trillion daily, doubling every few weeks. This growth stems from the demand unlocked after the engineering teams fully switched to Antigravity 2.0 + 3.5 Flash.

The more powerful Gemini 3.5 Pro was not released this time—it is expected to debut next month in 2026. Pichai revealed: "We are using it internally, and the results are great, but we are still polishing it."


2. Gemini Spark: Google's Personal Agent Finally Arrives

Spark is the most direct product answering how "Agents can reach beyond the 0.1%": a personal Agent running 24/7 in the cloud, connecting the entire Google ecosystem.

It runs on dedicated virtual machines in Google Cloud, powered by Gemini 3.5, using Google's internal Antigravity Harness. In other words, even when the laptop is closed or the phone is locked, it continues to run long-running tasks in the background.

The first batch of use cases Google defined for Spark is not science fiction: extracting facts from Gmail, Docs, Sheets, and Slides to write status reports; helping students update study guides; helping parents track school schedules; and monitoring inboxes for small businesses to prevent missing customer queries.

What truly matters is not these demos, but its positioning: Spark is Google's first step in pushing Agents out of the "chat window" and into "real digital life." It connects to Google's own tools first, with plans to open up to third-party MCPs later; the entry point is initially in the Gemini app, followed by Gmail, Google Chat, and Chrome, with Android Halo displaying execution progress on Android.

The rollout schedule is also clear: starting with trusted testers first, it will roll out next week as a beta to Google AI Ultra subscribers in the US. It pauses for confirmation before sensitive actions, allowing users to view the execution process and interrupt it at any time—this trust design is Google's prerequisite for encouraging everyday users to delegate tasks to an Agent.


3. Antigravity 2.0: Directly Facing Off Against Claude Code and Codex

Antigravity 1.0 was released about half a year ago (November 2025) and is currently used by "millions of developers." The key change in 2.0 is a shift in positioning—upgrading from a "coding environment" to a "multi-Agent orchestration platform."

Three-Tier Architecture:

  1. Desktop App: The central hub for interacting with Agents, capable of orchestrating multiple Agents simultaneously—one writing website code, one generating branding assets, and a third planning the product architecture, all working without conflict.
  2. Antigravity CLI: The command-line interface, built for developers who prefer to stay in the terminal.
  3. Antigravity SDK: The most critical layer—allowing applications to programmatically access the same Agent harness that drives Google's own products, optimized in tandem with Gemini models.

An Internal Case Study: Powered by the unreleased Gemini 3.5 Pro, 93 different Agents collaborated, consuming 2.6 billion tokens to build a complex system project completely from scratch.

In other words, Google is attempting to open up its internal Agent harness to developers in the form of an SDK. This is a direct face-off against Claude Code and Codex.

Pichai's assessment during discussions:

"The beauty of the developer community is that if you build something and it actually works, they will use it. Antigravity plus 3.5 Flash is a pretty magical experience, especially with how fast it is—3.5 Flash is already 4x faster than other frontier models, and in Antigravity, it's about 12x."

More compelling evidence comes from internal testing: running multi-Agent collaboration with 3.5 Flash + Antigravity 2.0, the Agents built a working operating system entirely from scratch, with API costs kept under $1,000.

Companion Release: Security Agent CodeMender

CodeMender uses Gemini's advanced reasoning to automatically find and fix critical code vulnerabilities, directly writing code to patch the security flaws, effectively taking over the manual patching process. Currently, a small number of experts are invited to test the API, with a broader release coming "very soon."

Google's logic: As Agents write more and more code, cybersecurity will become increasingly critical, and agentic workflows are the inevitable path forward in this domain.


4. Gemini App: Completely Redesigned + Price Cuts Across the Board

Over the past year, the monthly active users (MAUs) of the Gemini app grew from 400 million to 900 million, with daily request volume increasing sevenfold, covering 230 countries and 70 languages. Google defines it as the "most widely accessible AI assistant in the world."

This redesign of the Gemini app introduces a new design language called Neural Expressive: all animations, colors, typography, and haptic feedback have been revamped. Responses are no longer static walls of text, but rather dynamic pages that mix images, timelines, charts, and Gemini Live.

But the real shift is changing from a "Q&A app" into the front-end for Agents. Gemini Omni is integrated directly into the app, allowing users to upload photos or videos, apply templates, change styles, generate videos, and even create an AI avatar that looks and sounds just like them.

Daily Brief is the first out-of-the-box Agent: every morning, it organizes Gmail, calendars, and tasks for the day, prioritizes them, and provides recommendations for the next steps. Its predecessor was the experimental project CC, which launched on Google Labs in December 2025 and has now been incorporated into the main product line. This path illustrates Google's strategy: proving the concept in Labs, then scaling it up in the Gemini app.

Starting May 19, Neural Expressive is rolling out globally; Daily Brief is available to Google AI subscribers in the US.

Subscription Pricing Adjustments

Tier Price Change
Pro $19.99/month Unchanged
Ultra (New) $100/month New tier, targeted at users who want heavy Agent services but don't need top-tier features
Ultra (Top-tier) $200/month Reduced from $250

Google had only set the top-tier Ultra at $249.99 in May 2025, and in just one year, it has effectively reduced the price by 20% while adding a middle tier. The intention is clear: to widen the pathway for heavy Agent services like Spark, making them accessible to ordinary paying users.


5. Gemini Omni: Integrating Veo and Nano Banana into the Main Gemini Architecture

If 3.5 Flash is designed to "get Agents running," Gemini Omni is meant to "make the world model a reality." This section was presented by Demis Hassabis himself.

Google positions Omni as "a model capable of generating any output modality from any input," starting with video output and expanding to images and text in the future.

  • First Release: Gemini Omni Flash, available starting May 19 to Google AI subscribers
  • Entry Points: Gemini app, Google Flow, YouTube Shorts
  • Watermarking: All content generated or edited by Omni comes with built-in SynthID
  • Omni Pro: Coming soon

The Relationship Between Omni, Veo, and Nano Banana

According to Google's own description:

Omni is an evolution of the core Gemini architecture, extending Gemini's initial native multimodal design from input processing to output generation. Video generation is the entry point (closely related to Veo), and it brings the editing capabilities of Nano Banana into Veo, enabling natural video modifications similar to traditional editing. In essence, Omni is the main evolution of Gemini, while Veo and Nano Banana are capability components built onto it.

Nano Banana has generated over 50 billion images to date, and Omni now extends this capability to video. Use cases include generating video from text prompts, as well as uploading videos for style modification, character replacement, and temporally consistent continuous editing.


6. Search + Google Workspace: Agents Entering Every Corner of Google

The Search Box's Biggest Upgrade in 25 Years

AI Mode is now one year old—launched at I/O 2025, it now has over 1 billion monthly active users, with "query volume doubling every quarter since launch."

This time, AI Mode is upgraded to be powered by Gemini 3.5 Flash. The Google Search team's assessment is that users have started treating search as an entry point for longer, more specific, and cross-modal tasks rather than just typing a few keywords.

Consequently, Google redesigned its most core portal—the search box. The new smart search box expands to accommodate longer, more conversational queries; AI query suggestions replace traditional autocomplete; and text, images, files, videos, and Chrome tabs can all serve as inputs. AI Overviews and AI Mode are also integrated, allowing users to ask follow-up questions directly from an overview response.

More importantly, Agents are starting to grow directly out of the search box.

Information Agents turn search from a "one-off query" into "continuous monitoring in the background": users can create Agents that run 24/7 to monitor markets, real estate listings, sports, shopping, or any specific topic, delivering synthesized updates when triggers are met.

Generative UI and mini-apps within search represent another dimension: bringing the agentic coding capabilities of Antigravity and Gemini 3.5 Flash directly to search results. When a user asks a complex question, search doesn't just return answers; it can generate interactive charts, simulators, dashboards, and even reusable tools in real-time.

This is Google's most formidable distribution capability: while others must educate users to open a new Agent product, Google only needs to add a layer of capability to the search box.


7. Google's Full-Stack Play in the Escalating Agent Arena

For some time, Google was largely absent from the Agent competition. This I/O reflects its urgency—and its response is a full-stack layout:

  • Model Layer: 3.5 Flash leverages speed and pricing to lower the marginal cost of Agents to a level acceptable to enterprises; Omni pushes the world model a step further.
  • Product Layer:
    • Spark for everyone
    • Daily Brief for all AI subscribers
    • Information Agents and Generative UI bring Agents into Google's most widespread distribution channel—the search box
    • Antigravity 2.0 provides developers with a direct alternative to Claude Code
  • Commercial Layer: UCP, AP2, and Universal Cart bring Agents into real-world transaction scenarios
  • Infrastructure Layer: The eighth-generation TPU introduces a "dual-chip" division of labor for the first time—TPU 8T dedicated to training, and TPU 8I dedicated to inference. Paired with JAX and Pathways, training tasks can span multiple data centers, scaling globally to over 1 million TPUs—"the world's largest training cluster."

A Few Numbers That Anthropic and OpenAI Might Worry About

  • The user scale of Google's AI across its major products is unanimously in the billions
  • 5 products have 3 billion users each
  • Capex for this year is projected at $180–190 billion, which is 6x that of 2022 ($31 billion)
  • Monthly tokens processed across its products reach 3.2 quadrillion, which is 7x the 480 trillion processed at the time of I/O 2025

Summary

Connecting these dots, what Google truly wanted to prove this time is not just "we have Agents too," but: we have the capacity to make Agents the default entry point.

Only with fast enough models can marginal costs be brought down; only with Spark running in the cloud can Agents truly run independent of user devices; only with the Gemini app and search box providing daily access points can users adopt them without changing habits; and only with lower subscription pricing can heavy Agent services transition from a few tech enthusiasts to mainstream paying users.

The same logic applies to the developer tools battlefield. Antigravity 2.0 is not merely replicating Claude Code or Codex; it is shifting the competition from "whose code completion is smarter" to "who can offer an entire Agent harness that includes models, execution environments, tool calls, sub-agent orchestration, and distribution channels."

If this holds true, the unit of competition in the next round of developer tools will no longer be IDE plugins, but platform-level Agent orchestration systems. The battle between Claude Code, Codex, and Antigravity won't just be over a programmer's code editor, but over who will become the control console where developers keep a fleet of Agents running continuously.

Thus, this I/O is a clear signal: Google hasn't conceded the primary AI battlefield; it is choosing to engage using the strategy it knows best—bringing models, products, entry points, infrastructure, and commercial systems to bear all at once.

The Agent arena going forward will not just be about whose model is smarter, but also about who can deliver capabilities to the most people in a stable, affordable, and trustworthy manner. This is where Google's advantages are becoming clear again.

For everyday users, this is the best possible scenario: the more anxious top foundation model providers get, the more powerful, cheaper, and more accessible the tools become. Behind every "bombshell release," there is another company forced to deliver a better product at a lower price.

Comments
Share

Comments