memv: Open-Source Memory System for AI Agents

memv is an open-source memory system designed for AI agents with a unique approach to knowledge extraction. Unlike traditional memory systems that extract every fact and rely heavily on retrieval for organization, memv focuses only on storing prediction errors. It uses predict-calibrate extraction, where before extracting knowledge from a new interaction, it predicts what the episode should contain based on existing knowledge. Only facts that were unexpected are stored, as importance is derived from surprise rather than from initial large language model (LLM) scoring.
Key Details
- Bi-temporal Model: Each fact is tracked by both event and transaction times, allowing queries like "what did we know about this user in January?"
- Hybrid Retrieval: Utilizes vector similarity (sqlite-vec) combined with BM25 text search (FTS5) through Reciprocal Rank Fusion.
- Contradiction Handling: New facts automatically contradict and invalidate older conflicting ones, yet the full history is preserved.
- SQLite Default: Zero external dependencies - no need for Postgres, Redis, or Pinecone.
- Framework Agnostic: Works with LangGraph, CrewAI, AutoGen, LlamaIndex, or plain Python.
- MIT Licensed: Compatible with Python 3.13+ and utilizes asynchronous operations.
A sample setup using memv:
from memv import Memory
from memv.embeddings import OpenAIEmbedAdapter
from memv.llm import PydanticAIAdapter
memory = Memory(
db_path="memory.db",
embedding_client=OpenAIEmbedAdapter(),
llm_client=PydanticAIAdapter("openai:gpt-4o-mini"),
)
async with memory:
await memory.add_exchange(
user_id="user-123",
user_message="I just started at Anthropic as a researcher.",
assistant_message="Congrats! What's your focus area?",
)
await memory.process("user-123")
result = await memory.retrieve("What does the user do?", user_id="user-123")
The project is currently at an early stage (v0.1.0), and feedback is encouraged, especially concerning the extraction approach and potential useful integrations.
📖 Read the full source: r/LocalLLaMA
👀 See Also

SkyClaw: Rust-Based Autonomous AI Agent Runtime
SkyClaw is an autonomous AI agent runtime built in Rust with a 7.1 MB binary that idles at 14 MB RAM and starts in under one second. It operates on five engineering principles including autonomy, robustness, and brutal efficiency.

Giving Claude a Local LLM as an Assistant via MCP on Mac
A developer connects Claude to a local Qwen 2.5 Coder 14B via Ollama and MCP, creating a no-cost assistant for delegating tasks like text processing and handling large files.

ProofShot CLI Gives AI Coding Agents Browser Verification Capabilities
ProofShot is an open-source CLI tool that lets AI coding agents verify UI features by recording browser sessions, capturing screenshots, and collecting console errors. It works with any agent that can run shell commands and generates self-contained HTML reports for human review.

Stockade: A New Orchestration Tool for Claude Code with Channel Support and Security Layers
Stockade is an orchestration tool built around Anthropic's Agent SDK that provides channel-based session management, RBAC, and fine-grained permissions for AI agents. It addresses limitations in OpenClaw and NanoClaw by offering more control while maintaining security through containerization and credential proxies.