Engram v1.0.0: Persistent Memory for Local LLMs via Knowledge Graph

What Engram Does
Engram solves the problem of LLMs forgetting everything between sessions by providing persistent memory via a knowledge graph. Unlike vector databases that only find similar text, Engram understands relationships and can reason over them.
Core Features
- Knowledge graph with typed entities, relationships, and properties
- Hybrid search combining BM25 + vector similarity using Ollama/OpenAI embeddings or local ONNX
- Confidence lifecycle where facts strengthen with confirmation, weaken with time, and correct on contradiction
- Inference engine with forward/backward chaining that derives new facts from rules
- Built-in MCP server that works with Claude Code, Cursor, and Windsurf out of the box
- HTTP REST API with 25+ endpoints on port 3030
- Built-in web UI for graph exploration, search, and natural language queries
- Peer-to-peer mesh sync between instances with ed25519 authentication
- CORS enabled for any frontend integration
Technical Details
The entire system runs as an 8.3 MB binary with zero external dependencies. All data lives in a single .brain file that can be copied to back up or moved to migrate. No cloud, Docker, Python, or external database is required.
MCP Integration
MCP configuration is simple:
{
"mcpServers": {
"engram": {
"command": "engram",
"args": ["mcp", "/path/to/knowledge.brain"]
}
}
}The MCP server exposes these tools: engram_store, engram_relate, engram_query, engram_search, engram_prove, and engram_explain.
Quick Start Commands
engram create my.brain
engram store "PostgreSQL" my.brain
engram serve my.brainAfter running engram serve, the web UI is available at http://localhost:3030.
Availability
Engram is free for personal, research, and education use, with a commercial license available. The source and releases are on GitHub.
📖 Read the full source: r/LocalLLaMA
👀 See Also

Soul MCP Server Adds Persistent Memory and Safety for Local LLMs
Soul is an open-source MCP server that provides persistent memory across sessions for local LLMs with two commands: n2_boot at start and n2_work_end at end. It includes Ark safety features that block dangerous commands like rm -rf and DROP DATABASE at zero token cost, plus cloud storage configuration.

Off Grid: Utilizing Phone Hardware for Offline AI Applications
Off Grid is an open-source app that uses your phone's hardware for offline AI tasks like text generation and voice transcription.

Prompt-caching MCP plugin automatically reduces Claude API costs by identifying stable context
The prompt-caching MCP plugin automatically identifies stable parts of context like system prompts and tool definitions, then marks them for Anthropic's caching feature to reduce API costs by 80-92% in coding sessions.

AlphaCreek: An MCP Server That Chunks SEC Filings to Cut Token Usage by 85%
AlphaCreek is a free MCP connector for Claude that reduces token consumption by ~85% when working with SEC filings by first returning a table of contents, then fetching only the sections the agent requests.