Vivold Consulting

Google's Gemini Spark brings always-on, cloud-run agentic help to consumers

Key Insights

Google unveiled Gemini Spark, a personal AI agent inside the Gemini app that takes action on a user's behalf and runs 24/7 on dedicated Google Cloud virtual machines, so it works in the background without a laptop open. Powered by Gemini 3.5 and Google's Antigravity harness, it handles long-horizon tasks and connects to third-party tools via MCP. It began rolling out to trusted testers, with a beta reaching Google AI Ultra subscribers in the US shortly after.

Stay Updated

Get the latest insights delivered to your inbox

Agentic AI aimed squarely at consumers

After bringing agents to developers and enterprises, Google used I/O to push the idea to everyday users with Gemini Spark, a personal AI agent in the Gemini app that navigates your digital life and takes action under your direction.

How it works

- It runs on dedicated virtual machines on Google Cloud and operates 24/7, so tasks continue in the background without keeping a device open.
- It's powered by Gemini 3.5 and Google's Antigravity harness, which lets it carry out long-horizon tasks.
- It integrates with tools - starting with Google's own and, in the following weeks, third-party tools via MCP (the Model Context Protocol).
- You can work with it in the Gemini app and, soon, through email and chat; on Android, a new UI space called Android Halo will show live agent progress, and later in the summer Spark will operate inside Chrome as an agentic browser.

The rollout

Google began rolling Spark out to trusted testers that week, with a beta coming to Google AI Ultra subscribers in the US the following week. It's part of a broader agentic push across Google's products - including a "Daily Brief" agent that synthesizes inbox, calendar, and tasks into a morning digest, and agentic capabilities arriving in Search. Spark is Google's clearest statement yet that it wants the agent, running safely in the cloud, to become a default way consumers get things done.

Related Articles

An AWS knowledge-graph deployment turned 6-month research cycles into 3 weeks - and the blueprint transfers far beyond pharma

An AWS GraphRAG deployment in pharmaceutical research cut R&D cycles by 87% - initial discovery that took six months now closes in three weeks - by fusing siloed internal databases and public literature into one queryable knowledge graph on Amazon Neptune Analytics and Bedrock (running Claude). Every answer comes with verifiable citations and a mapped reasoning path, which is exactly what regulated industries need for compliance. The architecture is modular and, crucially, transferable: any enterprise drowning in fragmented legacy data can copy this pattern.

SpaceX, Anthropic, and OpenAI listings will out-value every US VC-backed exit since 2000 - reshaping vendor economics for everyone

The new NVCA-Pitchbook Venture Monitor dropped a stunning claim: the pending OpenAI and Anthropic IPOs, together with SpaceX's listing, will generate more value than every US VC-backed exit since 2000 combined. SpaceX is already public at $1.77 trillion, and with both AI labs pushing toward trillion-dollar debuts, the trio should land north of $4 trillion - against roughly $70 billion in total US IPO proceeds last year. For anyone buying AI services, the labs' shift to public-market scrutiny will reshape pricing, transparency, and vendor stability.

A 14-person open-source team just became the default way 8.9M developers run local AI - and a lever for slashing inference bills

Ollama, the open-source tool that lets developers run open-weight AI models on their own machines in minutes, raised a $65M Series B led by Theory Ventures ($88M total), revealing it now serves 8.9 million developers monthly and sits inside 85% of the Fortune 500 - with just 14 employees. Founders Jeff Morgan and Michael Chiang previously built Docker Desktop, and they're repeating the play: abstract away the hardware pain, then monetise a cloud tier priced on GPU time rather than tokens. The backdrop is the industry's loudest cost debate: every company with heavy inference bills is under existential pressure to shift routine workloads to open models.