Engramx v3.4: MCP Server + SQLite Knowledge Graph Cuts Claude Code Token Usage by 89%

Engramx v3.4 is an MCP server combined with a SQLite-based knowledge graph that intercepts file reads at the agent boundary. When Claude Code attempts to read a file that engram has indexed, the hook returns a structural summary instead of the raw content. The result: the same edit, same diff, but far fewer tokens consumed per round trip.
Key details
- Benchmark: Real 87-file codebase; aggregate token reduction 89.1%. Best-case file dropped from 18,820 tokens to 306. The benchmark script is
bench/real-world.ts— you can run it on any project you own. - IDE support: Works across 8 IDEs natively: Claude Code (hooks + official plugin in review), Cursor (MDC + MCP + VS Code extension on OpenVSX), Cline, Continue.dev, Aider, Windsurf, Zed, and OpenAI Codex CLI. One install, one graph, all tools benefit.
- Local-first: SQLite database lives at
.engram/graph.dbin your repo. Nothing leaves your machine. Apache 2.0 licensed. No account, no telemetry. - Install:
npm install -g engramxthencd ~/your-projectandengram setup. For Cursor, runcode --install-extension nickcirv.engram-vscode. - Tracking: The
engram costcommand shows token savings per project per week. After 24 hours of normal use, the digest displays real numbers. - Upcoming v4.0 “Mesh + Spine”: Releases May 25. Opt-in federation layer for sharing mistakes and ADRs across machines without sharing source. Phase 1 already merged: ed25519 identity, 14-category PII gate, 1007 tests.
The tool directly addresses hitting Claude Code Max 5x limits in under 2 hours on real work — a single complex prompt jumping the session counter from 21% to 100%.
📖 Read the full source: r/ClaudeAI
👀 See Also

Android CLI and Skills for AI Agent Development Workflows
Google released Android CLI with commands like android create and android sdk install, plus Android Skills GitHub repository with modular instruction sets. Internal benchmarks show 70% reduction in LLM token usage and 3x faster task completion.

Local Memory System for AI Coding Tools Extracts 2,600+ Facts from Conversation Logs
A developer built a local memory layer that ingests conversation logs from Claude Code, Factory.ai, and Codex CLI, extracts structured facts using a local LLM, and auto-injects context into new sessions. After months of use, it has indexed 13,000+ messages and extracted 2,600+ facts.

Claudetop: Real-Time Cost Monitoring for Claude Code Sessions
Claudetop is an htop-like tool that shows real-time spending, cache efficiency, and model comparisons for Claude Code sessions. It provides slash commands like /claudetop:stats and smart alerts for cost milestones and efficiency issues.

Claude-Code v2.1.111 adds Opus 4.7 xhigh effort, /ultrareview, and PowerShell tool
Claude-Code v2.1.111 introduces the Opus 4.7 xhigh effort level between high and max, adds the /ultrareview command for cloud-based multi-agent code reviews, and begins rolling out PowerShell tool support on Windows. The update also includes interactive /effort controls, auto theme matching, and numerous bug fixes.