Recall: Local Project Memory for Claude Code — No Tokens Spent on Summaries

Recall is a fully-local project memory plugin for Claude Code that solves the cold-start problem without spending any model tokens on summarization. It captures session transcripts into .recall/history.md and condenses them into a compact context.md (~1–2K tokens) using a classical Python summarizer — not an LLM call.
How It Works
- During a session: The
Stop/SessionEndhooks append new activity incrementally to.recall/history.md— only new turns, fully local. - At session start: The
SessionStarthook surfacescontext.mdand prompts Claude to confirm: resume from saved context? and keep logging this session?
Key Advantages
- Zero token spend on memory: Summarization is done locally by a deterministic algorithm, not by an API call. No API key or external model required.
- Privacy: Transcripts (code, paths, secrets) never leave your machine. Most memory tools pipe context to a model endpoint; Recall doesn't.
- Low friction: No
pip install, no local model to run, no key configure — works offline immediately on plugin load.
Output Files
Two files in .recall/:
history.md— append-only log of prompts, replies, files touched, commands run.context.md— overwritten summary containing: goal, summary, next steps/open threads, files touched, where you left off.
Comparison with Built-in Claude Code Memory
| Feature | CLAUDE.md | --continue / --resume | Recall |
|---|---|---|---|
| What | Hand-written notes & rules | Reloads a prior conversation | Auto-captured session log + local summary |
| Upkeep | Manual | None (you pick session) | None — written as you work |
| Holds | Instructions to follow | Full prior transcript | Goal, files, commands, where you left off, next steps |
| Cost to resume | Small | Large (replays full transcript) | ~1–2K tokens (compact digest) |
| Form | Markdown you edit | Local session state | Plaintext in .recall/ — diffable & shareable |
| Claude treats it as | Instructions | The conversation | Fenced untrusted reference data |
In short: CLAUDE.md is how I want you to work; Recall is here’s what we did last time and where we stopped — produced offline with zero model tokens spent.
📖 Read the full source: HN LLM Tools
👀 See Also

nex-life-logger: Local Activity Tracker for OpenClaw Agents
nex-life-logger is a background activity tracker that runs locally on your machine, giving OpenClaw agents memory of your computer activities. It tracks browser history, active windows, and YouTube transcripts, storing everything in a local SQLite database with no cloud data transmission.

Heren Godot MCP: Persistent WebSocket Daemon Cuts AI–Godot Interaction Latency to ~20ms
Heren is a new MCP server for Godot that keeps a lightweight WebSocket daemon alive, achieving ~20ms operations instead of waiting for full engine cold starts. It provides 15 tools for scene management, debugging, GPU‑accelerated screenshots, and automatic shutdown after 3 minutes of inactivity.

Building a Local Voice AI Assistant with SwiftUI and CSM-1B on Apple Silicon
A developer built mobiGlas, a SwiftUI app that pairs with OpenClaw to enable hands-free conversations via AirPods, using local voice cloning (CSM-1B on M2 Ultra) and no cloud APIs.

Throttle Meter: Open-Source Claude Code Usage Meter for macOS
Open-source macOS menu bar app that reads local Claude Code logs to show real-time 5-hour and weekly usage, with threshold notifications and token-saving hooks. Also has a €19 commercial sibling with Exact mode (reads claude.ai's internal API via Safari).