Vivold Consulting

Google rebuilds Search around background agents and on-the-fly generative interfaces

Key Insights

Google is extending Search into an agentic experience, introducing information agents - personalized AI agents that work in the background 24/7 to find what you need and help you act - rolling out over the summer to AI Pro and Ultra subscribers. It's also infusing Search with agentic coding, so Gemini 3.5 Flash and Antigravity build custom, generative interfaces (dynamic layouts, interactive visuals, even persistent dashboards) for individual queries, free for everyone. The move builds on AI Mode, which passed 1 billion monthly users in a year.

Stay Updated

Get the latest insights delivered to your inbox

Search becomes a set of agents working for you

Google laid out how Search changes in what it calls the agentic era, building on momentum from AI Mode - its biggest-ever Search upgrade, which surpassed 1 billion monthly active users within a year - and AI Overviews, now at over 2.5 billion.

What's new

- Information agents: personalized AI agents you set up to work in the background, around the clock, to surface what you need at the right moment and help you take action. They roll out over the summer, starting with Google AI Pro and Ultra subscribers.
- Generative UI: by infusing Search with agentic coding via Gemini 3.5 Flash and Antigravity, Search can build custom experiences for individual questions - dynamic layouts and interactive visuals - available to everyone this summer, free of charge.
- Custom dashboards: for longer-running tasks, Search can build persistent, returnable dashboards or trackers, almost like mini-apps for a specific task, with Antigravity-built experiences coming to Pro and Ultra subscribers in the US.

Why it matters

The framing is that Search is becoming less a series of individual queries and more an ongoing conversation that connects users to the web while doing more of the work for them. Coupled with agents that act and interfaces that assemble themselves per query, it's one of the more concrete pictures of how generative AI reshapes the product that remains Google's core - and the one that brings AI to more people than any other.

Related Articles

An AWS knowledge-graph deployment turned 6-month research cycles into 3 weeks - and the blueprint transfers far beyond pharma

An AWS GraphRAG deployment in pharmaceutical research cut R&D cycles by 87% - initial discovery that took six months now closes in three weeks - by fusing siloed internal databases and public literature into one queryable knowledge graph on Amazon Neptune Analytics and Bedrock (running Claude). Every answer comes with verifiable citations and a mapped reasoning path, which is exactly what regulated industries need for compliance. The architecture is modular and, crucially, transferable: any enterprise drowning in fragmented legacy data can copy this pattern.

SpaceX, Anthropic, and OpenAI listings will out-value every US VC-backed exit since 2000 - reshaping vendor economics for everyone

The new NVCA-Pitchbook Venture Monitor dropped a stunning claim: the pending OpenAI and Anthropic IPOs, together with SpaceX's listing, will generate more value than every US VC-backed exit since 2000 combined. SpaceX is already public at $1.77 trillion, and with both AI labs pushing toward trillion-dollar debuts, the trio should land north of $4 trillion - against roughly $70 billion in total US IPO proceeds last year. For anyone buying AI services, the labs' shift to public-market scrutiny will reshape pricing, transparency, and vendor stability.

A 14-person open-source team just became the default way 8.9M developers run local AI - and a lever for slashing inference bills

Ollama, the open-source tool that lets developers run open-weight AI models on their own machines in minutes, raised a $65M Series B led by Theory Ventures ($88M total), revealing it now serves 8.9 million developers monthly and sits inside 85% of the Fortune 500 - with just 14 employees. Founders Jeff Morgan and Michael Chiang previously built Docker Desktop, and they're repeating the play: abstract away the hardware pain, then monetise a cloud tier priced on GPU time rather than tokens. The backdrop is the industry's loudest cost debate: every company with heavy inference bills is under existential pressure to shift routine workloads to open models.