Audit of every doc touching the global-palace rollout against the running
fleet. Each correction below was verified against the filesystem or the host,
not against another doc:
- synlig-primary-runbook: the decommission `rm -rf ~/.mempalace` now carries a
STOP block. That tree holds the fleet palace *and* the only copy of the
bearer token every client authenticates with; the old "empty today" comment
stopped being true when the palace was seeded on 2026-08-14. Adds an ordered
safe decommission, and drops count-based join verification.
- phase-1-exposure-runbook: new S3.8, how to verify a flip actually took --
the procedure that until now existed only in an untracked handover file.
Three claims that fail independently (env var / curl / the palace-path
discriminator) plus an explicit list of checks that produce FALSE POSITIVES:
drawer counts (both sides were seeded from the same palace, and `status`
counts chunks not drawers), write-then-read through the same transport, and
the `mempalace` CLI -- which has no remote support at all, so post-flip it
reads the dead local archive and reports success.
- rfc-001: status Draft -> Phases 0-1 implemented. Records that the join was a
file-level copy, which SIDESTEPPED the S7.6 diary-dedup question rather than
answering it -- so S7.6 remains a hard blocker for the second machine, which
is the one that will actually exercise merge semantics.
- ARCHITECTURE, SKILL, contrib/README, extensions/pi/README all claimed pi
feeds the palace automatically, unconditionally. That is gated on
mempalace-toolkit >= 29e660e and every deployed image predates it, so the
claim is currently false fleet-wide. Each site now states the gate plus a
check that inspects the *deployed* file rather than repo HEAD.
- extensions/pi/README: plaintext http://mempalace.lan example -> https
endpoint; the two transports are either/or (no dual-write, no local mirror);
the bridge fails CLOSED, so "the agent has no mempalace_* tools" is the
expected symptom of a server/token/DNS fault, not of a broken install.
- contrib/README: documents mempalace-serve.service, which this directory has
shipped since day one without explaining it (linger, the load-bearing
172.17.0.1 bind and why loopback is the unsafe-looking-safe option, the
token path, and an uninstall warning).
- Fixes a pre-existing stray ```sh fence that was swallowing S3.2's heading and
the token command into a code block.
Docs only; no behaviour change.
pi ↔ MemPalace MCP bridge
The canonical source of ~/.pi/agent/extensions/mempalace.ts — the TypeScript
extension that wires MemPalace's MCP
server into the pi coding-agent
harness. Installs wake-up context injection, per-tool schema passthrough,
and a /mempalace-diary slash-command.
This directory only holds the bridge. Pi's own base config (keybindings,
environment loader, settings template) lives in the sibling
pi-toolkit repo — split out
2026-05-05 so opencode-devbox
can build slim containers that include pi without dragging in mempalace's
dependencies (~300 MB).
Jump to:
- What it does
- Transport: local vs external
- Automatic transcript feeding
- The
Type.Unsafegotcha - Deploying pi with mempalace on a new machine
- Fail-soft, identity, debugging
What it does
- Connects to MemPalace and does the MCP handshake (
initialize+notifications/initialized+tools/list). By default it spawnsmempalace-mcpas a local stdio subprocess (StdioMcpClient); if$MEMPALACE_REMOTE_URLis set it instead talks to a shared MemPalace over HTTP (RemoteMcpClient) and spawns no local process — see Transport. - Registers each MCP tool as a pi tool with its real
inputSchemapassed through viaType.Unsafe(...)(see gotcha below). - Wake-up auto-injection (
before_agent_start, one-shot per fresh session): callsmempalace_status+mempalace_diary_readand injects the result as amempalace-wakeupsystem message so the agent orients itself the way~/.agents/skills/mempalace/SKILL.mddescribes. Skipped on resume/fork (context is already in the thread). - Automatic transcript feeding (
session_shutdown, and a debouncedagent_settled): stages + mines this pi installation's own session transcripts into the palace with no user action needed — as of mempalace-toolkit29e660e(2026-08-12); see the version gate below, because "the extension is installed" does not imply "this copy can feed". Unlike the diary below, this needs no LLM turn — it's a subprocess + a tool call — so it can run onsession_shutdownwhere the diary cannot. See Automatic transcript feeding. - Manual wind-down via a
/mempalace-diary [topic]slash command: sends a prompt asking the LLM to callmempalace_diary_writewith an AAAK-formatted entry summarizing the session. This one stays manual because it needs the LLM to compose the entry, andsession_shutdownfires too late to drive another LLM turn — a constraint that applies to the diary specifically, not to feeding (see above).
Automatic transcript feeding
⚠️ Version gate — requires mempalace-toolkit ≥
29e660e(2026-08-12), and "installed" is not the same question as "capable". Feeding was added to this extension on 2026-08-12. A copy baked into a container image built before that date has no feed path at all — its entiresession_shutdownhandler isclient.stop()— and it fails the only way a memory system must not: silently, looking exactly like a healthy run with nothing to do.Check the deployed artifact, never the repo.
/opt/*in an image is baked at build time and can be days behind a bind-mounted clone, and~/.pi/agent/extensions/mempalace.tsis usually a symlink into that baked copy:grep -c MEMPALACE_FEED "$(readlink -f ~/.pi/agent/extensions/mempalace.ts)" # 0 = cannot feedZero hits means this machine needs the fallback recipes in
contrib/until it is rebuilt, regardless of what the toolkit repo's HEAD looks like. Date the deployed copy withstatplus that content probe — notgit log, which fails with "detected dubious ownership" inside a root-owned/opttree. As of 2026-08-14 the whole pi-devbox fleet fails this check.
The bridge feeds this pi installation's own session transcripts into the
palace by itself — no scheduler, no cron, no manual invocation. It fires on
session_shutdown (covers quit, /new, /resume, /fork) and on a
debounced agent_settled (covers a long session that later crashes, since a
hard kill runs no shutdown handler at all).
The work is split across two processes, and the reason is a hard constraint,
not a style choice: the palace is single-writer. A live pi session
always holds it through this extension's own mempalace-mcp subprocess, so
an unattended mempalace mine from anywhere else fails outright with
palace ... is held by PID <n>. The bridge therefore:
- Runs
mempalace-pi-session --prepare --reason <trigger> --wing <wing>as a subprocess. This does every palace-free step — parse pi's JSONL, apply the quality threshold, stage the export, and (remote mode only)rsyncit to the palace host — and prints one line,MINE_SOURCE=<path>, without ever touching the palace. - Calls the
mempalace_mineMCP tool through this extension's own client on that path. Going through the client that already holds the lock is the only way to write during a live session, and it automatically targets whichever palace the bridge is pointed at — local stdio or a shared remote one.
mempalace-pi-session (in this repo's bin/) is the actual exporter and
owns the quality gate, the remote transport, and every flag — see its
--help for the full reference; this section only covers the extension's
side of the wiring.
Env knobs (extension side):
| Var | Default | Effect |
|---|---|---|
MEMPALACE_FEED |
1 |
Set 0 to disable automatic feeding entirely. |
MEMPALACE_FEED_BIN |
mempalace-pi-session |
Helper to run. |
MEMPALACE_FEED_WING |
wing_conversations |
Target wing — passed to both the exporter and the mempalace_mine call. |
MEMPALACE_FEED_DEBOUNCE_MS |
600000 (10 min) |
Minimum gap between mid-session (agent_settled) feeds. Bounds crash loss to one window instead of a whole session. |
MEMPALACE_FEED_PREPARE_TIMEOUT_MS |
120000 |
Kills a wedged --prepare subprocess. |
MEMPALACE_FEED_MINE_TIMEOUT_MS |
30000 |
Caps the mempalace_mine call so a stalled palace can't hang session exit. |
Remote palace: if $MEMPALACE_REMOTE_URL is set (see
Transport), mempalace_mine's source path is
expanded on the server, which cannot see this machine's transcripts —
that's exactly why step 1 above rsyncs first in that mode. Configure the
inbox with MEMPALACE_PI_SSH_TARGET (required for remote feeding — feeding
is silently skipped without it), MEMPALACE_PI_SSH_CONFIG, and
MEMPALACE_PI_REMOTE_PATH; see mempalace-pi-session --help.
Concurrency: overlapping triggers coalesce — a session_shutdown landing
while a debounced tick is still running joins that in-flight feed instead of
racing it. mempalace-pi-session itself also takes a non-blocking flock,
so even two independent invocations (e.g. this extension and the
container-start catch-up some devbox images run) never race each other;
losing that race is harmless because the next trigger re-exports from
scratch.
Transport: local vs external
The bridge speaks the same MCP protocol over two interchangeable transports, chosen at load time:
-
Local (default) — spawns
mempalace-mcpas a stdio subprocess; the palace lives wherever that process opens it (default~/.mempalace). This is the hardened path with per-request timeouts and respawn/self-heal (below). -
External — set
MEMPALACE_REMOTE_URLto a MemPalace HTTP endpoint (e.g.https://mempalace.jordbo.se/mcp, the live fleet primary — full path including/mcp, no trailing slash) and the bridge connects over HTTP instead, spawning no local process. Use this to share one palace across several harnesses/containers (pi + opencode + native).MEMPALACE_REMOTE_TOKEN, if set, is sent asAuthorization: Bearer <token>. Usehttps://for anything crossing a network — the plaintexthttp://example that stood here until 2026-08-14 predated the reverse proxy.The two transports are either/or, decided once at load time: with the URL set, writes go only to the remote palace. There is no dual-write, no local mirror, and no local
mempalace-mcpprocess at all.⚠️ Consequence: once
MEMPALACE_REMOTE_URLis set, themempalaceCLI on that machine is no longer a valid way to inspect or feed the palace the agent is using. The CLI has no remote support whatsoever — its only selector is--palace <path>— so it reads and writes the LOCAL on-disk archive. After a flip that archive is frozen, yetmempalace status/mempalace searchstill report a plausible drawer count and look exactly like success: a false-positive machine. Memories filed with the CLI post-flip land in the dead archive, not in the shared palace. Use the agent's own palace tools (which go over HTTP), and mine backfills on the palace host.Serve such an endpoint with
mempalace serve --host 172.17.0.1 --port 8765(thepi-devbox/opencode-devboxrepos ship adocker-compose.mempalace.ymlfor exactly this).The HTTP transport is authenticated as of mempalace 3.6.0 — earlier docs here said otherwise, from the v1.3.0 era.
servemints a bearer token, keeps it 0600, passes it via the environment (never argv), compares it withhmac.compare_digest, and refuses to bind a non-loopback host without one unless--allow-insecure. It also pinsHostand allowlistsOrigin(anti-DNS-rebinding), and can terminate TLS itself.Two binds to avoid.
0.0.0.0publishes the palace to the whole LAN. And127.0.0.1is the trap that looks safe: the Host pin is enforced only on loopback binds, so behind a tunnel every proxied request 403s — and token auto-minting is gated on the bind being non-loopback, so it starts with no authentication at all, no warning. Bind the docker0 gateway (172.17.0.1): reachable from the host and its containers, not from the LAN. Seedocs/phase-1-exposure-runbook.md.Implementation note: the HTTP client (
RemoteMcpClient) is vendored frompi-extensions'mcp-loader.ts. AMCP-STREAMABLE-HTTP-CLIENT-SYNCtoken keeps the two copies from drifting —scripts/check-mcp-client-sync.shfails if they diverge (it skips gracefully when thepi-extensionscheckout isn't present).
Fail-soft
If mempalace-mcp can't be spawned (PATH missing, binary crashes at
startup, …) the extension logs to stderr and returns early. pi keeps
working without palace tools rather than refusing to start.
In remote mode the triggers differ but the outcome is identical. An
unreachable server, a DNS failure, or an HTTP 401 from a wrong/expired token
all end the same way: after bounded retries the extension prints
mempalace-mcp unavailable after retries; continuing without palace tools and
does not register the palace tools.
It is fail-closed, not fail-local: it does not quietly fall back to the
local palace, so a remote outage can never scatter memories into a local copy
nobody will look at again. The practical corollary, worth knowing before you
debug the wrong layer: "the agent has no mempalace_* tools" is the
expected symptom of a server, token, or DNS fault, not of a broken install.
Diagnose it with a direct curl to MEMPALACE_REMOTE_URL — see
docs/phase-1-exposure-runbook.md
§3.8. The design rationale for de-registering rather than degrading is in
docs/rfc-001-global-palace.md §2 and §4.1.
Identity
agent_name for diary calls comes from $MEMPALACE_AGENT_NAME, defaulting
to "pi". First diary write against that identity creates wing_<name>
in the palace. Set the env var if you want to run pi under a distinct
identity on a given machine (e.g. pi-laptop vs pi-server).
Stall protection (per-request timeout)
Every JSON-RPC request to mempalace-mcp carries a timeout. Without it, a
wedged server (classically: an OrbStack/virtiofs cold-open of a large
chroma.sqlite3 or an HNSW load) leaves the awaiting promise pending
forever, which freezes the pi TUI — ESC cancels the LLM stream, not a
pending tool execute(). On timeout the extension rejects the request
and kills the stalled child (SIGTERM→SIGKILL), so pi gets a clear
error instead of hanging. This is a per-REQUEST timeout, not a process-lifetime
one — the long-lived server is only killed when a request genuinely stalls.
MEMPALACE_MCP_TIMEOUT_MS— tool-call/request timeout. Default60000. Kept short on purpose: a query taking this long is genuinely wedged.MEMPALACE_MCP_INIT_TIMEOUT_MS—initialize+tools/listhandshake timeout. Default300000. Deliberately generous: a genuine first cold-open over virtiofs can legitimately take minutes, and killing a still-progressing init only to respawn and re-pay the same cold cost is strictly worse than waiting.- Set either to
0to disable (legacy unbounded behavior).
Self-heal (respawn instead of a permanent latch)
A stall-kill (or any crash) used to be a permanent latch: available
flipped off and stayed off until you restarted pi. It is now self-healing —
the next tool call transparently respawns mempalace-mcp and retries.
- Respawns use capped exponential backoff so a persistently-broken
server can't hot-loop:
MEMPALACE_MCP_MAX_RESPAWNSattempts (default2; set0to disable self-heal and keep the old fail-fast latch), withMEMPALACE_MCP_RESPAWN_BACKOFF_MS(default1000) doubled per attempt. - The budget resets on any successful JSON-RPC response — proof the server is actually live — so a server that recovers regains full patience, while one that keeps dying hits the cap and stays down (then restart pi).
- Why the long init timeout and bounded respawn compose rather than overlap: once a server has opened the palace once, the OS page cache is warm, so respawn cold-opens are fast. The long init timeout prevents killing a healthy first cold-open; the respawn handles a genuinely dead server cheaply afterwards. (Note the HNSW deserialize is CPU work that isn't page-cacheable across spawns, which is exactly why we can't rely on respawn-warming alone and keep the generous init budget.)
- The initial startup is tolerant too: if the very first
start()fails, the extension runs the same bounded respawn before falling back to fail-soft (pi keeps working without palace tools).
Debugging
MEMPALACE_EXT_DEBUG=1— surfacemempalace-mcpstderr into pi's stderr. Without this, stderr is drained silently so a misbehaving server doesn't flood the TUI.- If a tool call fails with a generic "Internal tool error", spawn
mempalace-mcpmanually with raw JSON-RPC on stdin to read the server-side error — much faster than guessing.
The Type.Unsafe gotcha
Earlier versions of this extension registered every MCP tool with
parameters: Type.Object({}, { additionalProperties: true }), which
discarded each tool's real inputSchema. The LLM then saw no parameter
names and had to guess, leading to bugs like mempalace_diary_read
being called with agent= instead of the required agent_name= and
crashing the Python server with TypeError: missing 1 required positional argument.
The fix (≈ lines 160-170) is to wrap the incoming JSON Schema with
Type.Unsafe<...>(tool.inputSchema). TypeBox schemas are plain JSON
Schema at runtime plus a Symbol marker, so wrapping an
externally-sourced schema with Unsafe is sufficient — no conversion
to a full TypeBox tree is needed, and the LLM now sees every tool's
real parameter names.
If you ever need to re-loosen the schema for debugging, fall back to
the Type.Object({}, { additionalProperties: true }) default only for
that specific tool, not globally.
Deploying pi with mempalace on a new machine
This is the "pi + memory" recipe. For pi without mempalace, see
pi-toolkit's README.
0. Prerequisites
- Shell: zsh + oh-my-zsh recommended (both toolkits install loaders into
~/.oh-my-zsh/custom/; bash works too, installers print the manualsourcesnippet). git,node≥ 20,uv,tmux≥ 3.2, pi installed upstream.- AWS credentials reachable via
AWS_PROFILE— only if usingamazon-bedrockas pi's provider.
1. Dotfiles (if you keep one)
Brings ~/.config/pi/.env (AWS creds, git-crypt encrypted), tmux CSI-u
extended keys, and other machine state:
git clone <your-dotfiles> ~/src/dotfiles
cd ~/src/dotfiles
git-crypt unlock <key>
./provision.sh --profile <profile> # or your equivalent tool
2. Install pi upstream
brew install pi-coding-agent # macOS
# or see https://github.com/earendil-works/pi for Linux
pi --help # creates ~/.pi/agent/
3. Install pi-toolkit (base pi config)
git clone ssh://git@gitea.jordbo.se:2222/joakimp/pi-toolkit.git ~/pi-toolkit
cd ~/pi-toolkit && ./install.sh
Symlinks keybindings.json, copies pi-env.zsh into
~/.oh-my-zsh/custom/, and prints the settings.json bootstrap command.
4. Bootstrap pi settings
cp ~/pi-toolkit/settings.example.json ~/.pi/agent/settings.json
$EDITOR ~/.pi/agent/settings.json # eu./us./anthropic: prefix
5. Install mempalace CLI + this toolkit
uv tool install mempalace
git clone ssh://git@gitea.jordbo.se:2222/joakimp/mempalace-toolkit.git ~/mempalace-toolkit
cd ~/mempalace-toolkit && ./install.sh
Detects pi, symlinks mempalace.ts into ~/.pi/agent/extensions/.
Also detects pi-toolkit artifacts and prints a green check (or a warning
telling you to install pi-toolkit first if you skipped step 3).
6. Register mempalace MCP with opencode (if applicable)
Skip if this box is pi-only. Otherwise:
- Install
opencode-toolkitso~/.config/opencode/.envis sourced into every shell (GitHub / Gitea / other MCP server tokens). - Register the mempalace MCP server in
~/.config/opencode/opencode.json— see root README § Registering mempalace with opencode.
7. First run
exec zsh
pi # should start with defaults; wake-up injection shows palace status
If the wake-up doesn't print, run MEMPALACE_EXT_DEBUG=1 pi to surface
mempalace-mcp stderr.
Verification checklist
# MCP bridge in place
ls -la ~/.pi/agent/extensions/mempalace.ts # → this repo
# pi-toolkit artifacts also in place
ls -la ~/.pi/agent/keybindings.json # → pi-toolkit
ls -la ~/.oh-my-zsh/custom/pi-env.zsh # cp from pi-toolkit
# Env loaded
zsh -ic 'echo $AWS_PROFILE $AWS_REGION'
# Palace reachable
mempalace status
Uninstall
cd ~/mempalace-toolkit && ./install.sh --uninstall --yes # bridge only
cd ~/pi-toolkit && ./install.sh --uninstall --yes # pi base config
# Leaves pi itself, mempalace CLI, and ~/.config/pi/.env alone.
File layout
mempalace-toolkit/
└── extensions/
└── pi/
├── README.md ← this file
└── mempalace.ts ← symlinked into ~/.pi/agent/extensions/
Pi base config (keybindings, env loader, settings template) lives in
pi-toolkit. install.sh
detects pi via ~/.pi/agent/extensions/ and runs a check_pi_toolkit
probe that warns if pi-toolkit's artifacts are missing.