Mnemos: an MCP server for persistent Claude Code memory

Claude Code forgets everything between sessions, forcing you to re-explain conventions, corrections, and context every time. Mnemos (GitHub) is an MCP server that fixes that by storing and replaying persistent memory across sessions.
How it works
- On session start, it pushes a ranked context block (conventions, corrections, skills, hot files, recent session summaries) into Claude's prompt.
- Records corrections as
tried / wrong_because / fix. Three corrections on the same topic auto-promote into a reusable skill withWhen this applies / Avoid / Dosections — deterministic pattern mining, no LLM in the loop. - Bi-temporal store: facts carry valid/invalid timestamps, so "we used to use X, now Y" works without stale context.
- Compaction recovery: one tool call restores the goal and key decisions after Claude Code compacts mid-session.
- Prompt-injection scanner at write boundary (instruction overrides, zero-width unicode, MCP spoofing).
- Retrospective replay: regenerate any past session as markdown with everything learned since layered in, paste it back to Claude, ask "what would I do differently now."
Stack & install
- Single static Go binary, 15 MB. No Python, no Docker, no vector DB, no CGO.
- SQLite + FTS5 retrieval, optional cosine similarity if Ollama is running.
- Install (MIT, free, no paid tier):
curl -fsSL https://raw.githubusercontent.com/polyxmedia/mnemos/main/scripts/install.sh | bash
mnemos init
mnemos initauto-wires Claude Code, Claude Desktop, Cursor, Windsurf, and Codex CLI. Restart your agent andmnemos_*tools show up.
Who it's for
Developers using Claude Code who are tired of re-teaching conventions every session and want reproducible, token-free memory.
📖 Read the full source: r/ClaudeAI
👀 See Also

Proactive Context-Rot Detection in Claude Code: A Feature Suggestion from r/ClaudeAI
A Reddit feature suggestion proposes that Claude Code proactively detect context rot and offer a structured task-scoped handoff, generating a handoff file and spawning a new session automatically.

Otterly: Route OpenClaw Through Your Claude Code Subscription
Otterly is a small npm package that exposes the local Claude CLI as an OpenAI-compatible HTTP server, letting you bill OpenClaw requests to your Claude Code subscription instead of pay-per-token API rates.

Clooks: A Persistent Hook Runtime for Claude Code
Clooks is a persistent HTTP daemon that handles Claude Code hook dispatch without process spawning, reducing latency from ~34.6ms to ~0.31ms per invocation. It includes automatic migration, LLM handlers with prompt templates, dependency resolution, and plugin packaging.

Librarian MCP: Local AI Server for Persistent Context with Documents
Librarian MCP is an open-source Model Context Protocol server that runs locally and connects to Jan, LM Studio, or Claude Desktop, enabling AI models to search and analyze document collections while maintaining full conversation context and data privacy.