Persistent memory for the Pi coding agent. LaPis gives Pi a local memory layer for decisions, bugfixes, patterns, indexed code, indexed docs, and session context.
It runs as one Pi extension plus one local Node.js backend. Storage is SQLite by default at ~/.pi/memory/memory.db; there are no cloud dependencies and no API keys.
A 30-second tour of the LaPis lifecycle. The GIF below autoplays inline (silent); click it to watch on YouTube with the voiceover.
Source composition lives at repo-media/html-video/lapis-slideshow/index.html β edit the prompts and re-render with npx hyperframes render.
LaPis is a modular monolith: one installable Pi extension with clear internal ownership between Pi adapters, CLI routing, feature services, and shared platform/storage code. The same backend also serves an MCP stdio server (lapis mcp) and a Claude Code hooks bridge (lapis claude-code install). The Pi extension calls the backend through in-process dispatch() when possible, with child-process fallback for streaming operations such as indexing.
All three architecture views are interactive HTML β clickable nodes, themeable (light/dark), and animated request paths with marching dashes on highlighted edges.
The full stack from Pi prompt to SQLite row. Chips animate 3 representative request lifecycles.
π Open the architecture overview β Save memory Β· Recall at start Β· Index code/docs
The dependency view. Highlights the layered architecture, peer relationships between feature services, and the trust-sync bridge (the only explicit memoryβcode link).
π Open the module boundaries diagram β Request flow Β· Feature peers Β· Trust bridge
How data moves through the four primary operations into one local SQLite store. The trust path is highlighted in violet to mark it as the only memoryβcode bridge.
π Open the memory lifecycle diagram β Write Β· Read Β· Index Β· Trust
For dependency rules and module ownership details, see docs/ARCHITECTURE.md and docs/MODULE_MAP.md.
pi install git:github.com/GeneGulanesJr/LaPisRestart Pi and memory auto-wires on session start. Use pi update --extensions to keep it up to date.
LaPis does not install npm dependencies at runtime. If you are running from a local clone or developing the extension, install dependencies explicitly:
npm installLaPis also integrates with the Claude Code CLI β same persistent memory, code guardrails, and session lifecycle as the Pi extension, wired through Claude Code's MCP + hooks config.
npx -y @genegulanesjr/lapis claude-code install
npx -y @genegulanesjr/lapis claude-code doctor # verify installThis writes .mcp.json (MCP tools) and .claude/settings.json (lifecycle hooks). On first use, approve the LaPis MCP server via /mcp in an interactive claude session.
Full setup, hook mapping, troubleshooting, and install flags: docs/CLAUDE_CODE.md.
LaPis also integrates with Hermes Agent β same persistent memory, code guardrails, and trust tracking as the Pi extension, wired through Hermes' MCP client and shell-hook system.
npx -y @genegulanesjr/lapis hermes install
npx -y @genegulanesjr/lapis hermes doctor # verify installThis writes mcp_servers.lapis + hooks: into $HERMES_HOME/config.yaml, first-use consent, and a Hermes skill. Restart Hermes (or /reload-mcp) to load the MCP tools; hooks load at process start.
Full setup, hook mapping, troubleshooting, and install flags: docs/HERMES.md.
For MCP hosts that need LaPis tools without hooks:
npx -y @genegulanesjr/lapis mcpSee docs/MCP.md.
-
Remembers across sessions - decisions, bugfixes, patterns, discoveries, and constraints persist.
-
Auto-injects context - new sessions start with relevant memories loaded.
-
Indexes code - web-tree-sitter parses JS/TS/TSX/Go/Python/Rust/SQL for semantic code lookup and analysis.
-
Indexes docs - Markdown sections, links, glossary terms, and code examples become searchable.
-
Tracks trust - memories linked to changed code lose confidence; stable linked code recovers trust.
-
Deduplicates memory - similar saves are merged or flagged before they clutter recall.
-
Manages workspaces - project isolation is explicit through create/list/archive workflows.
-
Cleans stale memory - the Dream Cycle removes superseded, never-useful, and replaced memories based on quality signals.
-
Exposes an HTTP API - optional REST server for programmatic access to missions, milestones, working units, todo/ledger domain, and code analysis.
-
Pre-coding intelligence - preflight checks combine memory, code, and docs into before-coding context.
-
Compresses CLI output - token-saving output compression reduces context window usage automatically.
-
Memory dashboard - observability command for memory health, statistics, and index quality.
The compact wire format (src/platform/protocol/compact-format.js) uses compact encoding to reduce the token footprint of analysis responses inside Pi's context window. The benchmark runs real CLI commands against indexed repos, passes output through compactResponse(), and compares byte sizes.
Run it with:
node bench/bench-tokens.jsLatest local run: May 24, 2026, with fresh reindexes for both repos. call-hierarchy and blast-radius were skipped because the benchmark could not select a representative symbol with callers.
| Tool | LaPis / PiMemoryExtension | PCBuilder |
|---|---|---|
| importance | 21% | 24% |
| hotspots | 0% | 0% |
| dead-code | 44% | 50% |
| coupling | 37% | 40% |
| extraction | 23% | 22% |
| import-graph | 18% | 20% |
| cycles | 0% | 0% |
| overall | 40% | 49% |
| LaPis / PiMemoryExtension | PCBuilder | |
|---|---|---|
| Repo size | 292 files / 6,913 symbols | 171 files / 207,599 symbols |
| Raw JSON | 692.6 KB | 36.7 MB |
| Compact format | 417.5 KB | 18.5 MB |
| Bytes saved | 275.1 KB | 18.1 MB |
| Tokens saved | ~80,500 tokens | ~5,436,007 tokens |
See bench/README.md for benchmark usage and interpretation notes.
The paired benchmark measures whether memory helps by running the same task twice: once with LaPis disabled and once with LaPis active. It is an internal regression and directional benchmark, not a comprehensive external evaluation. It is designed to catch regressions in LaPis behavior; it should not be read as a claim about typical real-world usage, where prompts, repositories, model behavior, provider cache state, and tool choices vary.
Run it with:
npm run bench:pi-pairedLatest run: 2026-05-24 (results are written locally under bench/results/, which is gitignored, so the path below is not present in a fresh clone) β bench/results/pi-paired-2026-05-24T14-47-20-651Z/report.json
| Metric | Memory On | Without Memory | Savings |
|---|---|---|---|
| Facts correct | 18/18 | 18/18 | no loss |
| Active tokens | 3,192 | 42,954 | -92.6% |
| Wall time | ~86s | ~233s | -63.1% |
| Tool calls | 6 | 49 | -87.8% |
| Failed tools | 0 | 4 | -100% |
| Category | Facts (on) | Tokens (on) | Savings |
|---|---|---|---|
| prior-decision | 3/3 | 128 | 99.0% |
| bug-history | 3/3 | 603 | 94.2% |
| staleness | 3/3 | 426 | 94.2% |
| navigation | 3/3 | 72 | 99.0% |
| negative-control | 6/6 | 1,963 | 65.7% |
| Category | What it tests |
|---|---|
| prior-decision | Recalls an architectural decision and its rationale, then names the current module involved. |
| bug-history | Recalls why a fix exists, including the historical failure mode that is not obvious from the final code alone. |
| staleness | Checks whether LaPis warns that an indexed code view may be stale and should be verified or reindexed before trust. |
| navigation | Uses memory to jump to the likely hook/module and confirm where extension wiring lives. |
| negative-control | Asks current-source questions that should not need memory facts; memory-on should route cheaply to code lookup instead of adding overhead. |
Memory-on achieved perfect accuracy with 92.6% fewer active tokens overall. Memory-dependent tasks saved 94.2-99.0% active tokens in this run.
The paired benchmark also reports behavior counters. In the latest run, memory-on used 6 total tools, 4 code-oriented tools, 4 memory tools, 12 assistant turns, and 0 failed tools. These counters help distinguish real memory regressions from normal provider cache and latency variance. The negative-control tasks are current-source questions; they should avoid memory facts and route quickly through memory-code search plus targeted reads when code verification is needed.
- Node.js 20+
better-sqlite3for local SQLite access- No Python dependency
- No API keys or cloud services
-
docs/INDEX.md- documentation map and integration transports. -
CONTRIBUTING.md- contributor workflow and checks. -
docs/ARCHITECTURE.md- architecture overview and dependency rules. -
docs/MODULE_MAP.md- module ownership and entry points. -
docs/ARCHITECTURE_MODULARIZATION.md- detailed modularization assessment. -
docs/COMMANDS.md- command reference. -
docs/API.md- HTTP API and CLI reference. -
docs/CONFIGURATION.md- config file and stored data. -
docs/CLAUDE_CODE.md- Claude Code CLI integration (MCP + hooks). -
docs/HERMES.md- Hermes Agent integration (MCP + hooks + skill). -
docs/MCP.md- standalone MCP server (tools only). -
docs/DREAM_CYCLE.md- stale-memory cleanup behavior. -
docs/TUTORIAL.md- step-by-step usage guide. -
docs/GITHUB_ISSUE_BREAKDOWN.md- modularization issue breakdown. -
docs/code-indexing.md- async code indexing. -
docs/SKILL.md- extension skill overview.
ISC




