Recall: Local Project Memory for Claude Code — No Tokens Spent on Summaries

Recall is a fully-local project memory plugin for Claude Code that solves the cold-start problem without spending any model tokens on summarization. It captures session transcripts into .recall/history.md and condenses them into a compact context.md (~1–2K tokens) using a classical Python summarizer — not an LLM call.
How It Works
- During a session: The
Stop/SessionEndhooks append new activity incrementally to.recall/history.md— only new turns, fully local. - At session start: The
SessionStarthook surfacescontext.mdand prompts Claude to confirm: resume from saved context? and keep logging this session?
Key Advantages
- Zero token spend on memory: Summarization is done locally by a deterministic algorithm, not by an API call. No API key or external model required.
- Privacy: Transcripts (code, paths, secrets) never leave your machine. Most memory tools pipe context to a model endpoint; Recall doesn't.
- Low friction: No
pip install, no local model to run, no key configure — works offline immediately on plugin load.
Output Files
Two files in .recall/:
history.md— append-only log of prompts, replies, files touched, commands run.context.md— overwritten summary containing: goal, summary, next steps/open threads, files touched, where you left off.
Comparison with Built-in Claude Code Memory
| Feature | CLAUDE.md | --continue / --resume | Recall |
|---|---|---|---|
| What | Hand-written notes & rules | Reloads a prior conversation | Auto-captured session log + local summary |
| Upkeep | Manual | None (you pick session) | None — written as you work |
| Holds | Instructions to follow | Full prior transcript | Goal, files, commands, where you left off, next steps |
| Cost to resume | Small | Large (replays full transcript) | ~1–2K tokens (compact digest) |
| Form | Markdown you edit | Local session state | Plaintext in .recall/ — diffable & shareable |
| Claude treats it as | Instructions | The conversation | Fenced untrusted reference data |
In short: CLAUDE.md is how I want you to work; Recall is here’s what we did last time and where we stopped — produced offline with zero model tokens spent.
📖 Read the full source: HN LLM Tools
👀 See Also

Local voice-to-text transcription for OpenClaw using Parakeet TDT 0.6b v3
A developer has converted NVIDIA's Parakeet TDT 0.6b v3 model to run locally via ONNX on CPU, supporting 25 European languages. The model provides an OpenAI-compatible API endpoint through a Docker container, allowing integration with OpenClaw for audio file transcription.

SwiftUI Agent Skill: Enhancing View Development with AI
SwiftUI Agent Skill is an open-source tool that uses AI to improve SwiftUI view development by embedding best practices and optimizations.

OmniCoder-9B: 9B Parameter Coding Agent Fine-Tuned on 425K Agentic Trajectories
Tesslate released OmniCoder-9B, a 9-billion parameter coding agent model fine-tuned on Qwen3.5-9B's hybrid architecture. It was trained on 425,000+ curated agentic coding trajectories from Claude Opus 4.6, GPT-5.4, GPT-5.3-Codex, and Gemini 3.1 Pro.

LocalSynapse MCP Server Adds macOS Support and Search Improvements
LocalSynapse, an offline MCP server for searching local documents, now supports macOS and includes fixes for multi-word search queries. The developer has implemented feedback-driven improvements including position-adjusted click boosting and time decay as promotion.