ScreenMind: Local-First AI Memory That Indexes Your Entire Computer Activity

ScreenMind is a local-first AI memory system that continuously captures your screen, transcribes meetings, and indexes voice notes, building a persistent, searchable timeline of everything you do on your computer. It uses perceptual hashing to only trigger when content changes, then runs each frame through Gemma 4 E2B via llama.cpp for vision analysis, chat, and audio processing.
Key Features
- Screen capture with perceptual hashing — only stores frames when content actually changes
- Searchable timeline — query past activity: "that error message from earlier," "what was I working on at 3pm?"
- Chat with your history — persistent AI context from your entire session
- Meeting transcription — auto-detects Zoom, Teams, and Google Meet
- Voice memos — processed via Gemma 4's audio encoder
- Natural language automations — write them in plain English Markdown
- MCP integration — connect to Claude and Cursor
Technical Stack
- Models: Gemma 4 E2B (handles vision, chat, audio)
- Backend: Python + FastAPI
- Storage: SQLite
- Inference: llama.cpp with Q4 quantization
- Hardware: 4GB+ VRAM
The author notes that GPU scheduling between vision, chat, and audio tasks is the main inference optimization challenge. The project is still workflow-driven rather than fully autonomous — retrieval quality and onboarding friction are areas needing improvement.
GitHub: ayushh0110/ScreenMind
📖 Read the full source: r/LocalLLaMA
👀 See Also

Contextium: Open-Source Persistent Context Framework for Claude Code
Contextium is a structured git repo framework that provides persistent context for Claude Code sessions, using a CLAUDE.md file as a context router to lazy-load relevant markdown files. The open-source version includes a template with 6 sample apps and 27 integration docs.

Fullerenes: Open-source persistent memory layer for coding agents cuts tokens by 64% on SWE-bench
Fullerenes uses a local SQLite knowledge graph built via Tree-sitter to give coding agents like Claude Code persistent memory, reducing token usage by 64% on SWE-bench and up to 96.6% on internal benchmarks.

Cross-Model Review Loop for AI Coding Agents Catches Critical Planning Flaws
A developer built a cross-model review system where a second AI model reviews plans from coding agents before execution, catching critical flaws like rollback failures and security holes. The tool is MIT licensed and includes a TUI dashboard.

Savecraft MCP Server Provides Claude with Accurate Magic: The Gathering Data
Savecraft is an open-source MCP server that parses MTG Arena Player.log locally, syncs game state, and gives Claude access to 12 expert reference modules built on real Magic: The Gathering data. The tool prevents Claude from hallucinating card names and rules by providing access to actual Arena data, draft recommendations from 17Lands, and the complete Scryfall database.