skill-depot: A Local-First Memory and Skill System for MCP-Compatible AI Agents

✍️ OpenClawRadar📅 Published: March 27, 2026🔗 Source
skill-depot: A Local-First Memory and Skill System for MCP-Compatible AI Agents
Ad

What skill-depot Does

skill-depot addresses the problem of AI agent skills and knowledge piling up across scattered directories. Instead of loading everything into context (wasting tokens) or loading nothing (forgetting learned material), it provides a retrieval system that stores agent knowledge as Markdown files and uses vector embeddings to semantically search and selectively load only what's relevant.

How It Works

Agents interact with skill-depot through three levels of detail:

  • skill_search("query") returns search results with name, score, and snippet
  • skill_preview("skill-name") returns a structured overview with headings and first sentence per section
  • skill_read("skill-name") returns the full Markdown content

The skill_learn tool allows agents to create or append knowledge on the fly, returning actions like "created" or "appended" with tags merged.

Technical Implementation

  • Embeddings: Uses local transformer model all-MiniLM-L6-v2 via ONNX (384-dim vectors, ~80MB one-time download)
  • Storage: SQLite + sqlite-vec for vector search
  • Fallback: BM25 term-frequency search when the model isn't available
  • Protocol: MCP with 9 tools (search, preview, read, learn, save, update, delete, reindex, list)
  • Format: Standard Markdown + YAML frontmatter (same format Claude Code and Codex use)
Ad

Setup and Use Case

Setup is simple: npx skill-depot init. The tool is designed for local-first, zero-config, MCP-native use with no API keys to manage, no server to run, and no framework lock-in. The tradeoff is a narrower scope—it doesn't do session management or automatic memory extraction (yet).

Comparison with Other Tools

  • mem0: Good for managed memory layer with polished API, but has cloud dependency
  • OpenViking: Full context database with session management, multi-type memory, and automatic extraction from conversations
  • LangChain/LlamaIndex memory modules: Solid if already in those ecosystems

Future Considerations

The developer is considering adding:

  • Memory types (distinguishing between skills, memories, and resources)
  • Deduplication to detect near-duplicate entries
  • TTL/expiration for temporary knowledge auto-cleanup
  • Confidence scoring where memories reinforced across multiple sessions rank higher

📖 Read the full source: r/openclaw

Ad

👀 See Also

Tendr Skill: Deterministic CLI Operations for Agent Memory Management
Tools

Tendr Skill: Deterministic CLI Operations for Agent Memory Management

Tendr Skill is an Agent Skill that separates reasoning from execution for structured long-term memory, allowing agents to decide what needs changing while a CLI tool handles structural operations deterministically. It supports [[wikilinks]] and explicit semantic hierarchies across files.

OpenClawRadar
Tredict MCP Server Enables Claude to Create and Push Training Plans to Sports Watches
Tools

Tredict MCP Server Enables Claude to Create and Push Training Plans to Sports Watches

A developer built a Tredict MCP Server for Claude.ai and Claude Code that creates complex endurance training plans via prompts and automatically uploads structured workouts to Garmin, Coros, Suunto, and Wahoo watches. The server includes an MCP App for visual feedback within Claude chat.

OpenClawRadar
Exploring LiveDocs: An AI-native Data Analysis Notebook
Tools

Exploring LiveDocs: An AI-native Data Analysis Notebook

LiveDocs offers a reactive notebook environment allowing data teams to perform multi-step analyses and maintain analysis end-to-end with the help of an AI agent.

OpenClawRadar
Scaling Karpathy's Autoresearch with 16 GPUs: Results and Methods
Tools

Scaling Karpathy's Autoresearch with 16 GPUs: Results and Methods

The SkyPilot team gave Claude Code access to 16 GPUs on a Kubernetes cluster to run Karpathy's Autoresearch project. Over 8 hours, the agent submitted ~910 experiments, reduced validation bits per byte from 1.003 to 0.974 (2.87% improvement), and reached the best validation loss 9x faster than sequential execution.

OpenClawRadar