Vivold Consulting

ChatGPT's new 'Dreaming' memory system fights staleness and scales to free users

Key Insights

OpenAI is rolling out its most capable ChatGPT memory system yet, built on a background process called dreaming that synthesizes memories for freshness, accuracy, and scale. It targets the staleness and correctness problems that show up across hundreds of millions of users and multi-year histories. It's live for Plus and Pro in the US now, with Free and Go users following over the coming weeks after a roughly 5x cut in serving cost.

Stay Updated

Get the latest insights delivered to your inbox

ChatGPT's memory grows up

OpenAI is rolling out what it calls its most capable memory architecture yet for ChatGPT - an upgrade aimed squarely at the three problems that plague memory at scale: it goes stale, it gets things wrong, and it's expensive to run across hundreds of millions of users and multi-year time horizons.

From sticky notes to "dreaming"

The feature has evolved in stages, and the framing is useful:

- Saved memories (April 2024) only recorded what you explicitly asked it to remember, which in practice felt like talking to someone who jotted a few notes and forgot everything else - and those notes went stale over time.
- Dreaming (first introduced April 2025) added a background process that automatically curates memory by drawing on your chat history, so context that comes up naturally gets retained without an explicit "remember this."
- The new release is a more capable, compute-efficient architecture built on top of dreaming - referred to as Dreaming V3 - that finally stands on its own rather than just supplementing saved memories.

Crucially, what the system synthesizes is reviewable: a memory summary page lets you see the highlights of what ChatGPT knows, correct or add details, and tell it which topics to raise and when.

What "good memory" is supposed to do

OpenAI frames quality around three objectives, each of which it says the new system improves on internal evals:

- Carry context forward so you don't reintroduce yourself every chat - for instance, asking for camera gear compatible with "my setup" and getting answers tuned to gear you discussed months ago.
- Follow preferences and constraints, whether explicit ("I'm vegetarian"), instructional ("don't bring that up again"), or implicit ("I live near San Francisco").
- Stay current over time, so "I'm going to Singapore in July" becomes "went to Singapore in July 2026" once the trip passes, and recommendations snap back to your home location.

Why the scale story matters

The headline operational win is cost: recent improvements cut the compute needed to serve dreaming to Free users by roughly 5x, which is what makes it practical to extend to everyone and to raise memory capacity for paying tiers. The bigger point is strategic - dreaming now gives OpenAI a shared memory foundation across all users, and personalization that deepens over time is central to making ChatGPT stickier and more useful. Memory controls and an FAQ accompany the release.

Related Articles

Discovery Loop aims to automate science itself - and Google is funding the startup draining its own bench, as Hassabis exits the DeepMind CEO role

Jeff Dean, Google's chief scientist and 30th employee, is leaving after 27 years to found Discovery Loop, a public benefit corporation using AI to automate scientific research - taking co-founders Sanjay Ghemawat, Quoc Le (Google Brain), and Oriol Vinyals (DeepMind) with him. Google is a founding investor and cloud partner, supplying compute for at least the first year, with Radical Ventures and Khosla Ventures co-leading the seed. In the same announcement, Demis Hassabis steps down as DeepMind CEO to become chairman and Alphabet chief scientist, with Koray Kavukcuoglu taking over Gemini model development. Alphabet stock fell about 4%.

Abbott orders audits of every new project as ERCOT's queue hits 474GW, roughly 90% of it data centres

Governor Greg Abbott announced that all new Texas data-centre projects must be audited by the Public Utility Commission and grid operator ERCOT - a sharp turn for a state whose loose regulation and cheap power made it second only to Virginia for data centres. The trigger is a staggering queue: ERCOT's interconnection requests doubled from 233GW in January to 474GW, about 90% data centres, more than five times the grid's all-time peak demand. Audits will demand power and water use, noise mitigation, light controls, tax-incentive use, and ownership details - after a voluntary survey that most operators simply ignored.

Volta and Bitdeer will build a 133MW Nvidia Vera Rubin data centre in Norway - Anthropic's latest move in a compute land grab

Anthropic has reportedly signed a $10 billion, six-year compute deal with Volta, an AI cloud startup founded only earlier this year, per Bloomberg. Volta is partnering with crypto-mining firm Bitdeer to develop the data centre - located in Norway, delivering 133 megawatts, and running Nvidia's Vera Rubin architecture - and is a member of Nvidia's Cloud Partner programme. It caps an aggressive capacity spree that also includes recent compute deals with SpaceX and Amazon, as Anthropic races rivals for the scarcest input in the industry.