Skip to content

feat(persistence): generation persistence — client snapshot + durable media bytes#987

Open
AlemTuzlak wants to merge 37 commits into
feat/persistence-corefrom
feat/generation-persistence
Open

feat(persistence): generation persistence — client snapshot + durable media bytes#987
AlemTuzlak wants to merge 37 commits into
feat/persistence-corefrom
feat/generation-persistence

Conversation

@AlemTuzlak

@AlemTuzlak AlemTuzlak commented Jul 23, 2026

Copy link
Copy Markdown
Contributor

Generation Persistence

Stacked on #984 (feat/persistence-core) — please review/merge that first. Rebased onto its current tip.

Generation persistence in two halves:

1. Client resume snapshot (read-only)

As a media generation streams, the client builds a lightweight GenerationResumeSnapshot (run identity, status, errors, result metadata + artifact refs — never the bytes) and writes it to a persistence storage adapter. The option reuses the shared ChatStorageAdapter contract, so localStoragePersistence / sessionStoragePersistence / indexedDBPersistence work with no type argument:

const snapshots = localStoragePersistence({ keyPrefix: 'my-app:generation:' })
useGenerateImage({ connection, persistence: snapshots })

Hooks expose resumeSnapshot / resumeState / pendingArtifacts / resultArtifacts across react/solid/vue/svelte/angular. No resume() action; reconnect to an in-flight stream is the delivery layer's job (resumable streams, #955, merged), wired via a durability adapter + GET handler exactly as for chat.

2. Durable media-byte storage (server, opt-in)

When the persistence backend provides both an artifacts (ArtifactStore) and a blobs (BlobStore) store, withGenerationPersistence writes each generated file's bytes to the blob store (key artifacts/<runId>/<artifactId>), records an ArtifactRecord, attaches PersistedArtifactRefs to the result, and emits generation:artifacts so the client records them. memoryPersistence() ships both stores; any backend implementing the two contracts works. Extraction is customizable via extractArtifacts / nameArtifact.

Changes

  • @tanstack/ai — result-transform machinery (resultTransforms / artifactInputs on GenerationMiddlewareContext, applyGenerationResultTransforms), threadId/runId options on the image/audio/speech/transcription activities, and generation:artifacts emission from streamGenerationResult.
  • @tanstack/ai-utilsbase64ToUint8Array.
  • @tanstack/ai-persistenceArtifactStore + BlobStore contracts, in-memory impls in memoryPersistence(), and byte-persistence in withGenerationPersistence (extractArtifacts/nameArtifact).
  • @tanstack/ai-client + framework hooks (react/solid/vue/svelte/angular) — the client snapshot + persistence option + artifact refs.
  • @tanstack/ai-event-client — optional threadId/runId on generation events.
  • Docsdocs/persistence/generation-persistence.md (client snapshot, byte storage + serve route, reconnect), kiira-verified.
  • Changeset — minor across the affected packages.

Verification

  • Builds green across ai-utils, ai, ai-event-client, ai-client, ai-persistence.
  • test:types0 errors in ai / ai-client / ai-persistence.
  • Tests — ai-persistence 78 pass (incl. 12 artifact + the chat resume suites), ai-client generation 50 pass, ai core stream/middleware/activity 92 pass.
  • kiira — doc passes (3 snippets).

Scope note

Durable SQL/R2 artifact+blob backends (Drizzle/Prisma/Cloudflare R2) are not in this PR: memoryPersistence plus any custom ArtifactStore+BlobStore cover the contract today, and the durable backends (which add SQL schema + migrations per package) are a clean follow-up. generateVideo's own job-polling artifact path is likewise deferred (image/audio/speech/transcription persist bytes).

AlemTuzlak and others added 19 commits July 22, 2026 18:06
…kends

Server-side persistence for chat(): durable thread messages, run records, and
interrupts via the withChatPersistence middleware, with pluggable backends.

- @tanstack/ai-persistence: store contracts, withChatPersistence /
  withGenerationPersistence middleware, memoryPersistence reference store,
  conformance testkit. Locks (LockStore/InMemoryLockStore/LocksCapability) live
  here rather than core; the sandbox-consumer bridge is deferred.
- -drizzle / -prisma / -cloudflare: backend store implementations + migration /
  schema / models CLIs. Cloudflare D1 delegates to the drizzle backend.

Reconciled against the shipped ephemeral-interrupt engine: the middleware
records interrupts and gates new input, and delegates resume-tool-state
reconstruction to the engine (resume batch + interrupt bindings in history).

Claude-Session: https://claude.ai/code/session_01RqjWdHxvmMrhjbd8dYvENp
The persistence adapter now stores one combined { messages, resume? } record
per chat id, so a full page reload restores the transcript, rehydrates pending
interrupts, and rejoins an in-flight run through joinRun when the connection is
durability-backed. Legacy bare-array records are still read.

Adds localStoragePersistence / sessionStoragePersistence / indexedDBPersistence
(+ StorageUnavailableError and the ChatPersistedState / ChatStorageAdapter
types). Durability rides the existing option, so every framework integration
gets it with no framework-specific code.

Claude-Session: https://claude.ai/code/session_01RqjWdHxvmMrhjbd8dYvENp
New /persistent-chat route: useChat with localStoragePersistence on the client
and withChatPersistence(sqlitePersistence) on the server, so a full page reload
restores the conversation on both ends. Adds a nav link and README section.

Claude-Session: https://claude.ai/code/session_01RqjWdHxvmMrhjbd8dYvENp
…ility

New docs/persistence section: overview, chat-persistence, browser-refresh,
controls, custom-stores, sql-backends, drizzle, prisma, cloudflare, migrations,
internals. Wires the nav and updates the client chat/persistence page for the
combined { messages, resume } record and built-in storage adapters.

Claude-Session: https://claude.ai/code/session_01RqjWdHxvmMrhjbd8dYvENp
Update the agent skills for the new surface: withChatPersistence server
middleware and its backends, and the client browser-refresh durability
(combined persistence record, storage adapters, joinRun rejoin, all frameworks).

Claude-Session: https://claude.ai/code/session_01RqjWdHxvmMrhjbd8dYvENp
Provider-free durable harness route + client page + spec proving message
restore after reload and interrupt-survives-reload via localStorage. Mid-stream
joinRun rejoin is covered by ai-client unit tests and delivery-durability.

Claude-Session: https://claude.ai/code/session_01RqjWdHxvmMrhjbd8dYvENp
Add the object form `persistence: { store, messages?: boolean }`. `messages:
false` caches only the tiny resume pointer, keeping large transcripts off the
client while durability rejoin and interrupt restore still work and the server
stays authoritative for history. A bare adapter remains shorthand for
`{ store, messages: true }`, so this is backward compatible and every framework
passthrough is unchanged.

The persistent-chat example gains a history branch on its GET route
(loadThread by threadId, distinct from the per-run delivery replay) so a
server-authoritative reload can hydrate the transcript. Docs + chat-experience
skill document the lever.

Claude-Session: https://claude.ai/code/session_01RqjWdHxvmMrhjbd8dYvENp
Rewrite the persistence overview into a concept + decision page: the three
problems (dropped stream, lost-on-reload, no durable record), the two
independent layers (delivery durability vs state persistence), client vs server
halves, the reload/rehydration timeline, and a when-to-pick-each guide. Add a
back-link from the resumable-streams overview so the two sections cross-reference.

Claude-Session: https://claude.ai/code/session_01RqjWdHxvmMrhjbd8dYvENp
…tence

localStoragePersistence / sessionStoragePersistence / indexedDBPersistence now
default their type parameter to ChatPersistedState and to a JSON codec, so
`persistence: localStoragePersistence()` needs no type argument and no
serialize/deserialize pair. Drops the IsJsonSerializable type gate that forced a
codec for the chat record (UIMessage already round-trips as JSON on the wire).

Client persistence now keys on `threadId` (the conversation identity), so a
reload with the same threadId restores the same record; `id` becomes an optional
storage-key override. The storage adapters and persistence types are re-exported
from every framework package, so a single import from @tanstack/ai-react (etc.)
works.

Claude-Session: https://claude.ai/code/session_01RqjWdHxvmMrhjbd8dYvENp
reconstructChat(persistence, request) returns a thread's stored messages as a
JSON Response, so a server-authoritative client can hydrate its transcript on
load from a one-line GET handler instead of hand-rolling loadThread + Response.
…curate resume

Add a "What we recommend" section to the overview: client resume-pointer-only
plus server persistence plus one GET that rehydrates history and resumes durable
streams, with the reasoning. Update every snippet to the zero-config
localStoragePersistence() and threadId, use reconstructChat for history, and
discriminate the resume GET with durability.resumeFrom() instead of sniffing
query params (the run id rides the X-Run-Id header, the offset the Last-Event-ID
header). Example, e2e page, and chat-experience skill match.
Rename browser-refresh to client-persistence and make it the single home for
the client story: turning it on, what a reload restores, the two cache modes
(everything vs resume-pointer-only) with when to use each, and the three storage
backends with when to use each. Remove the legacy docs/chat/persistence page
(client content now lives in the persistence section) and repoint its links.

Make the other persistence docs server-only: drop the client rows from the
controls decision table and the browser-storage section from internals, leaving
a pointer to the client guide. The overview stays the cross-cutting map. Update
all cross-links, the chat-experience skill source, and the reconstructChat doc
reference.
In `{ messages: false }` mode a prior session's persisted record is
`{ messages: [], resume }`. The constructor treated that empty transcript as
authoritative and clobbered host-provided `initialMessages`, and the async
hydrate path applied `[]` on top, so a server-authoritative reload dropped the
history the app had fetched from the server.

The persisted transcript is now adopted only when the client actually caches it
(`cachesMessages`); in messages:false mode the client keeps `initialMessages`
and takes only the resume pointer from storage. This makes the recommended
server-authoritative flow work: on a mid-stream reload the app seeds history via
initialMessages (the reconstruct GET) while the client separately rejoins the
live run via joinRun (the resume GET), and the replayed run merges into the
seeded history by message id. Adds a test covering both together.

Also document in the overview that history hydration and run rejoin are two
separate GET requests, so the handler's if/else routes each and neither blocks
the other.
The primary chat persistence middleware is now `withPersistence`.
`withGenerationPersistence` is unchanged. Unreleased, so no alias is kept.
Updates all call sites, docs, skills, the example, and the changeset.

fix(ai-client): rejoin an in-flight run from an async persistence store

Auto-rejoin was gated on the synchronous read, so an async store
(indexedDBPersistence) restored messages and interrupts on reload but never
rejoined a mid-stream run. A guarded maybeRejoinInFlight now fires from both the
sync read and the async hydrate path; it rejoins a run at most once and never
while another run is already active (a fresh send wins). Adds a test that a
run rejoins from an async (Promise-returning) adapter.
@coderabbitai

coderabbitai Bot commented Jul 23, 2026

Copy link
Copy Markdown
Contributor

Important

Review skipped

Auto reviews are disabled on base/target branches other than the default branch.

Please check the settings in the CodeRabbit UI or the .coderabbit.yaml file in this repository. To trigger a single review, invoke the @coderabbitai review command.

⚙️ Run configuration

Configuration used: defaults

Review profile: CHILL

Plan: Pro Plus

Run ID: 21ef07e1-2cac-46df-81f9-5d36d57a641d

You can disable this status message by setting the reviews.review_status to false in the CodeRabbit configuration file.

Use the checkbox below for a quick retry:

  • 🔍 Trigger review
✨ Finishing Touches
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch feat/generation-persistence

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

@github-actions

github-actions Bot commented Jul 23, 2026

Copy link
Copy Markdown
Contributor

🚀 Changeset Version Preview

15 package(s) bumped directly, 38 bumped as dependents.

🟥 Major bumps

Package Version Reason
@tanstack/ai-angular 0.3.1 → 1.0.0 Changeset
@tanstack/ai-durable-stream 0.0.0 → 1.0.0 Changeset
@tanstack/ai-persistence 0.0.0 → 1.0.0 Changeset
@tanstack/ai-persistence-cloudflare 0.0.0 → 1.0.0 Changeset
@tanstack/ai-persistence-drizzle 0.0.0 → 1.0.0 Changeset
@tanstack/ai-persistence-prisma 0.0.0 → 1.0.0 Changeset
@tanstack/ai-preact 0.11.1 → 1.0.0 Changeset
@tanstack/ai-react 0.18.1 → 1.0.0 Changeset
@tanstack/ai-solid 0.15.1 → 1.0.0 Changeset
@tanstack/ai-svelte 0.15.1 → 1.0.0 Changeset
@tanstack/ai-vue 0.15.1 → 1.0.0 Changeset
@tanstack/ai-acp 0.2.3 → 1.0.0 Dependent
@tanstack/ai-anthropic 0.16.3 → 1.0.0 Dependent
@tanstack/ai-bedrock 0.1.4 → 1.0.0 Dependent
@tanstack/ai-claude-code 0.2.3 → 1.0.0 Dependent
@tanstack/ai-code-mode 0.3.8 → 1.0.0 Dependent
@tanstack/ai-code-mode-skills 0.3.11 → 1.0.0 Dependent
@tanstack/ai-codex 0.2.3 → 1.0.0 Dependent
@tanstack/ai-elevenlabs 0.2.34 → 1.0.0 Dependent
@tanstack/ai-fal 0.9.12 → 1.0.0 Dependent
@tanstack/ai-gemini 0.20.1 → 1.0.0 Dependent
@tanstack/ai-grok 0.14.9 → 1.0.0 Dependent
@tanstack/ai-grok-build 0.2.3 → 1.0.0 Dependent
@tanstack/ai-groq 0.5.3 → 1.0.0 Dependent
@tanstack/ai-isolate-node 0.1.47 → 1.0.0 Dependent
@tanstack/ai-isolate-quickjs 0.1.47 → 1.0.0 Dependent
@tanstack/ai-mistral 0.2.3 → 1.0.0 Dependent
@tanstack/ai-ollama 0.8.16 → 1.0.0 Dependent
@tanstack/ai-openai 0.17.1 → 1.0.0 Dependent
@tanstack/ai-opencode 0.2.3 → 1.0.0 Dependent
@tanstack/ai-openrouter 0.15.10 → 1.0.0 Dependent
@tanstack/ai-react-ui 0.8.15 → 1.0.0 Dependent
@tanstack/ai-sandbox 0.2.4 → 1.0.0 Dependent
@tanstack/ai-sandbox-cloudflare 0.2.4 → 1.0.0 Dependent
@tanstack/ai-sandbox-daytona 0.2.0 → 1.0.0 Dependent
@tanstack/ai-sandbox-docker 0.2.0 → 1.0.0 Dependent
@tanstack/ai-sandbox-local-process 0.2.0 → 1.0.0 Dependent
@tanstack/ai-sandbox-sprites 0.2.1 → 1.0.0 Dependent
@tanstack/ai-sandbox-vercel 0.2.0 → 1.0.0 Dependent
@tanstack/ai-solid-ui 0.7.14 → 1.0.0 Dependent
@tanstack/openai-base 0.9.9 → 1.0.0 Dependent

🟨 Minor bumps

Package Version Reason
@tanstack/ai 0.42.0 → 0.43.0 Changeset
@tanstack/ai-client 0.22.1 → 0.23.0 Changeset
@tanstack/ai-event-client 0.6.8 → 0.7.0 Changeset
@tanstack/ai-utils 0.3.1 → 0.4.0 Changeset

🟩 Patch bumps

Package Version Reason
@tanstack/ai-devtools-core 0.4.24 → 0.4.25 Dependent
@tanstack/ai-isolate-cloudflare 0.2.38 → 0.2.39 Dependent
@tanstack/ai-mcp 0.2.5 → 0.2.6 Dependent
@tanstack/ai-vue-ui 0.2.34 → 0.2.35 Dependent
@tanstack/preact-ai-devtools 0.1.67 → 0.1.68 Dependent
@tanstack/react-ai-devtools 0.2.67 → 0.2.68 Dependent
@tanstack/solid-ai-devtools 0.2.67 → 0.2.68 Dependent
ag-ui 0.0.2 → 0.0.3 Dependent

@nx-cloud

nx-cloud Bot commented Jul 23, 2026

Copy link
Copy Markdown

View your CI Pipeline Execution ↗ for commit 41397c1

Command Status Duration Result
nx run-many --targets=build --exclude=examples/... ✅ Succeeded 3s View ↗

☁️ Nx Cloud last updated this comment at 2026-07-23 15:46:54 UTC

A stale local build had masked real breakage: the ported persistence layer was
behind the current core. Reconciled against a clean build:

- withGenerationPersistence keys runs on `requestId` (the current
  GenerationMiddlewareContext has no runId/threadId).
- Restore server-authoritative resume: withPersistence translates persisted
  interrupts into resumeToolState and clears `config.resume`, so the engine
  skips its ephemeral (client-history) reconstruction, which the empty-messages
  persistence flow can't satisfy. (Reverts an incorrect earlier removal.)
- ai-client: never persist an empty record (no messages, no resume), so a
  cleared conversation is removed, not left as `{ messages: [] }` — fixes the
  clear/suppression regressions from the combined-record change.
- Tests updated to the combined-record shape, LockStore imported from
  @tanstack/ai-persistence, generation-context mocks and event shapes aligned to
  the current core.

The two-phase approval->client-tool continuation test is skipped with a TODO:
its exact resume-execution semantics depend on the engine and need reconciling
with the engine owner; single-phase approval and client-tool resume are covered.
@pkg-pr-new

pkg-pr-new Bot commented Jul 23, 2026

Copy link
Copy Markdown

Open in StackBlitz

@tanstack/ai

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai@987

@tanstack/ai-acp

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-acp@987

@tanstack/ai-angular

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-angular@987

@tanstack/ai-anthropic

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-anthropic@987

@tanstack/ai-bedrock

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-bedrock@987

@tanstack/ai-claude-code

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-claude-code@987

@tanstack/ai-client

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-client@987

@tanstack/ai-code-mode

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-code-mode@987

@tanstack/ai-code-mode-skills

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-code-mode-skills@987

@tanstack/ai-codex

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-codex@987

@tanstack/ai-devtools-core

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-devtools-core@987

@tanstack/ai-durable-stream

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-durable-stream@987

@tanstack/ai-elevenlabs

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-elevenlabs@987

@tanstack/ai-event-client

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-event-client@987

@tanstack/ai-fal

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-fal@987

@tanstack/ai-gemini

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-gemini@987

@tanstack/ai-grok

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-grok@987

@tanstack/ai-grok-build

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-grok-build@987

@tanstack/ai-groq

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-groq@987

@tanstack/ai-isolate-cloudflare

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-isolate-cloudflare@987

@tanstack/ai-isolate-node

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-isolate-node@987

@tanstack/ai-isolate-quickjs

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-isolate-quickjs@987

@tanstack/ai-mcp

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-mcp@987

@tanstack/ai-mistral

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-mistral@987

@tanstack/ai-ollama

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-ollama@987

@tanstack/ai-openai

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-openai@987

@tanstack/ai-opencode

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-opencode@987

@tanstack/ai-openrouter

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-openrouter@987

@tanstack/ai-persistence

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-persistence@987

@tanstack/ai-persistence-cloudflare

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-persistence-cloudflare@987

@tanstack/ai-persistence-drizzle

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-persistence-drizzle@987

@tanstack/ai-persistence-prisma

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-persistence-prisma@987

@tanstack/ai-preact

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-preact@987

@tanstack/ai-react

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-react@987

@tanstack/ai-react-ui

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-react-ui@987

@tanstack/ai-sandbox

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-sandbox@987

@tanstack/ai-sandbox-cloudflare

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-sandbox-cloudflare@987

@tanstack/ai-sandbox-daytona

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-sandbox-daytona@987

@tanstack/ai-sandbox-docker

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-sandbox-docker@987

@tanstack/ai-sandbox-local-process

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-sandbox-local-process@987

@tanstack/ai-sandbox-sprites

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-sandbox-sprites@987

@tanstack/ai-sandbox-vercel

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-sandbox-vercel@987

@tanstack/ai-solid

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-solid@987

@tanstack/ai-solid-ui

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-solid-ui@987

@tanstack/ai-svelte

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-svelte@987

@tanstack/ai-utils

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-utils@987

@tanstack/ai-vue

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-vue@987

@tanstack/ai-vue-ui

npm i https://pkg.pr.new/TanStack/ai/@tanstack/ai-vue-ui@987

@tanstack/openai-base

npm i https://pkg.pr.new/TanStack/ai/@tanstack/openai-base@987

@tanstack/preact-ai-devtools

npm i https://pkg.pr.new/TanStack/ai/@tanstack/preact-ai-devtools@987

@tanstack/react-ai-devtools

npm i https://pkg.pr.new/TanStack/ai/@tanstack/react-ai-devtools@987

@tanstack/solid-ai-devtools

npm i https://pkg.pr.new/TanStack/ai/@tanstack/solid-ai-devtools@987

commit: 41397c1

The persisted-state resume path already works end-to-end: withPersistence
rehydrates the paused thread into config.messages, so the engine reprocesses
the pending tool call from server state (not the omitted client history).
Approving a client tool therefore advances straight to the client-execution
interrupt without re-invoking the model, and feeding the client output drives
one final model call. The prior skip assumed a model re-invocation that does
not happen; assertions now match the real engine behavior.
…authoritative persistence

Switch the demo to the recommended setup: `persistence: { store, messages: false }`
so the client caches only the resume pointer and the server (SQLite) owns history.
A route loader hydrates the transcript from a server function that reads the
stored thread — SSR-safe (no relative fetch), sharing one lazily-opened store
with the API route.
Type `prismaPersistence`'s client argument structurally (`PrismaClientLike`)
instead of importing `PrismaClient` from `@prisma/client`. The runtime was
already structural and the delegate query API is identical across majors, so
this accepts a client from either the v6 `prisma-client-js` generator or the v7
`prisma-client` generator (emitted to a custom output, not `@prisma/client`).
Docs note both versions.
withPersistence.onFinish saved ctx.messages, but the chat engine only appends
an assistant turn to the middleware message list when it carries tool calls —
a run's terminal text reply is never appended. So a stored thread dropped the
assistant's final answer and a server-authoritative reload showed only user
messages. Reattach the terminal reply from info.content (the last turn's
accumulated text), guarded against duplication. Strengthen the unit test to
assert the assistant reply is stored (it previously only checked length > 0,
which masked this).
Add getWeather / rollDice server tools so the demo exercises the agent loop and
tool-call persistence (tool calls + results are stored and rehydrated on
reload), and replace the bare inline styles with a self-contained dark chat UI
that renders tool-call cards (input/output) alongside message bubbles, plus
suggestion chips and auto-scroll. Regenerate routeTree.gen.ts to register the
persistent-chat routes.
Layer a lightweight, read-only resume snapshot onto media generation.
As a run streams, the client builds a GenerationResumeSnapshot (run
identity, status, errors, result metadata + artifact refs — never media
bytes) and writes it to an optional GenerationServerPersistence store.

- ai-client: GenerationResumeSnapshot types + updateGenerationResumeSnapshot
  reducer; GenerationClient/VideoGenerationClient observe chunks, persist
  snapshots (serialized queue, warn-not-throw), expose getResumeSnapshot();
  disposed guard. No resume() action (stream re-attach is PR #955).
- ai-event-client: optional threadId/runId on generation events.
- react/solid/vue/svelte/angular hooks: persistence + initialResumeSnapshot
  options; expose resumeSnapshot/resumeState (+ pending/result artifacts).
- example: Persisted mode on the image generation route.
- docs: persistence/generation-persistence.md + nav entry.

Pairs with the existing withGenerationPersistence server middleware.
Drop the bespoke `GenerationServerPersistence` type and the `{ server }`
option wrapper. The `persistence` option is now a bare storage adapter
reusing the shared `ChatStorageAdapter` contract (aliased as
`GenerationPersistence`), so `localStoragePersistence` /
`sessionStoragePersistence` / `indexedDBPersistence` work for generations
exactly as they do for chat — matching main's ergonomics.
Default `localStoragePersistence` / `sessionStoragePersistence` /
`indexedDBPersistence` to a value-agnostic `TValue` so a bare, unannotated
call works for BOTH chat and generation persistence — the consuming
`persistence` option constrains the stored value. Generation docs/example now
use `localStoragePersistence({ keyPrefix })` with no type declaration.
PR #955 (resumable streams) is merged, so delivery durability is available
today — it was wrongly described as an unlanded future feature. Rewrite the
generation-persistence doc: the server example now wires a durability adapter
+ GET handler, and the delivery section explains that a dropped mid-generation
connection re-attaches through the same adapters useChat uses. Clarify that the
read-only snapshot carries run state (incl. runId) across reloads, while
hooks do not auto-resume on mount.
@AlemTuzlak
AlemTuzlak force-pushed the feat/generation-persistence branch from c1cd561 to a77d800 Compare July 23, 2026 14:38
Server-side artifact + blob storage for generated media, layered on
withGenerationPersistence.

- @tanstack/ai: result-transform machinery (resultTransforms/artifactInputs
  on GenerationMiddlewareContext, applyGenerationResultTransforms), threadId/
  runId on the image/audio/speech/transcription activities, and
  generation:artifacts emission from streamGenerationResult.
- @tanstack/ai-utils: base64ToUint8Array.
- @tanstack/ai-persistence: ArtifactStore + BlobStore contracts, in-memory
  impls in memoryPersistence(), and withGenerationPersistence byte-persistence
  (writes bytes to blobs, records ArtifactRecord, attaches PersistedArtifactRef,
  emits generation:artifacts) with extractArtifacts/nameArtifact options.
…+ blobs, serve route, extractArtifacts/nameArtifact)
@AlemTuzlak
AlemTuzlak force-pushed the feat/generation-persistence branch from 2f56980 to f3d9fe6 Compare July 23, 2026 14:45
@AlemTuzlak AlemTuzlak changed the title feat(persistence): client-side generation persistence feat(persistence): generation persistence — client snapshot + durable media bytes Jul 23, 2026
Add retrieveArtifact(persistence, id) and retrieveBlob(persistence, idOrRecord)
so a serve handler fetches a persisted generation artifact's metadata and bytes
without hand-rolling the blob key. artifactBlobKey is the shared key builder used
by both withGenerationPersistence (write) and retrieveBlob (read).
@tombeckenham
tombeckenham force-pushed the feat/persistence-core branch from 0b17612 to d2d08e8 Compare July 24, 2026 01:31
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant