skip to content
The Weighted Average

Agentic Engineering

Claude's Shared Memory Kills the Re-Brief Tax

Anthropic merged Claude chat and Cowork memory. At Opus 5 rates, re-explaining a 200k-token project context costs $1.00 per agent session you no longer pay.

A hand pulls a card from a library's wooden card catalog
A hand pulls a card from a library's wooden card catalog. Photograph by Ilya Semenov

Anthropic merged the memory systems behind Claude chat and Claude Cowork on Tuesday, so context learned across 200,000 tokens of one surface now carries into the other. What Claude retains is exposed as editable topic files that users can read or delete, and memories now update mid-conversation rather than at session end, TechCrunch reported from the announcement. The feature is on by default for Free, Pro, and Max across web, desktop, and mobile; for Team and Enterprise, memory stays off until an admin enables it. Strip away the convenience framing and this is a pricing change disguised as a product update.

Put a number on the re-brief

Every operator running long-lived agents pays a context tax: the tokens spent reconstructing what the model already knew in a different window. Price it. Claude’s Pro and Max plans carry a 200,000-token context window, and Opus 5 API input runs $5 per million tokens. Refilling a full context window therefore costs $1.00 per session in input tokens alone — before the agent does any work, and before output.

That figure is deliberately an upper bound; most re-briefs are a fraction of a full window. But it is the right bound to reason with, because agentic sessions trend toward filling context, not toward brevity. Anthropic’s June 2026 Economic Index report documents the asymmetry precisely: the median chat or Cowork conversation producing an article involves 13 rounds of back-and-forth, while the median Claude Code session producing the same artifact contains a single human prompt. Delegation rises when the model already has context; it collapses when the human has to rebuild it. Shared memory is a bet that removing re-briefing shifts users toward the one-prompt mode — which, per the same report, is also the mode that consumes 2.07 times the tokens of low-wage-mapped work.

Read those two facts together and the economics of this update invert nicely for Anthropic. Users save the dollar of re-brief input; Anthropic gains sessions that run longer, delegate more, and burn more output tokens at $25 per million on Opus 5. Nobody is being fleeced — the trade is genuinely good for a team whose agents keep forgetting the project — but it is not charity, and your monthly bill will not fall.

The governance surface is the real change

For anyone deploying Claude inside an organisation, the operational story is the defaults, not the dollars. Memory now writes across surfaces automatically on consumer plans, and Anthropic excludes sensitive categories by default: health data, race, ethnicity, religious beliefs, politics, and gender identity, with an opt-in toggle to include them and notifications when a sensitive topic is stored. Some categories are never stored — government IDs, Social Security numbers, criminal history, immigration status, and anything violating the usage policy.

Anthropic’s earlier memory launch established the pattern this update extends: project-scoped memories, an editable memory summary, and Incognito chats that never persist. The unification loosens those project boundaries in exchange for continuity, which is exactly the tradeoff a security reviewer should be asked to sign off on. The New Stack notes the memory files are organised by topic and can be deleted but not directly edited — you change them by asking Claude to change them, which is a strange audit posture for a record that now spans every surface a user touches.

There is a competitive edge to the announcement as well. An Anthropic spokesperson told The New Stack that “Claude also remains ad-free, so nothing in memory is used for ad targeting” — a pointed contrast with rivals monetising attention, and a line that will appear in enterprise procurement decks within the week. Whether it survives contact with Anthropic’s own revenue pressure is a separate question; the company is not immune to the same economics that have pushed its priciest tier to fight for share of its customers’ own spend.

Three actions follow for teams. Confirm your plan’s default: Enterprise and Team memory is off until an admin turns it on, so nothing has silently changed for governed deployments — but individual staff on personal Pro accounts are now accumulating cross-surface memory of work context. Decide whether the sensitive-topics toggle stays off as policy rather than as an accident of default. And treat the memory summary as a reviewable artifact, because a persistent, model-written record of what your organisation is working on is a new object in your data map.

The counterpoint worth holding: memory quality is not the same as memory presence. A benchmark published this month found that how a system renders retained history to the model changes results by up to 72.6 points across 500 questions and nine models, according to the RENDER paper — the artifact matters more than the retrieval. Anthropic has not published accuracy figures for the unified system, so the saved dollar is measurable while the quality is not. The verdict: enable it, instrument what it stores, and hold your evals on the outputs rather than the feature. This paper reached a similar conclusion about agent skill libraries degrading as they grow past twenty entries, and today’s lead makes the hardware version of the same argument about efficiency claims that only bind at a particular operating point.

Sources