memv: Open-Source Memory System for AI Agents

memv is an open-source memory system designed for AI agents with a unique approach to knowledge extraction. Unlike traditional memory systems that extract every fact and rely heavily on retrieval for organization, memv focuses only on storing prediction errors. It uses predict-calibrate extraction, where before extracting knowledge from a new interaction, it predicts what the episode should contain based on existing knowledge. Only facts that were unexpected are stored, as importance is derived from surprise rather than from initial large language model (LLM) scoring.
Key Details
- Bi-temporal Model: Each fact is tracked by both event and transaction times, allowing queries like "what did we know about this user in January?"
- Hybrid Retrieval: Utilizes vector similarity (sqlite-vec) combined with BM25 text search (FTS5) through Reciprocal Rank Fusion.
- Contradiction Handling: New facts automatically contradict and invalidate older conflicting ones, yet the full history is preserved.
- SQLite Default: Zero external dependencies - no need for Postgres, Redis, or Pinecone.
- Framework Agnostic: Works with LangGraph, CrewAI, AutoGen, LlamaIndex, or plain Python.
- MIT Licensed: Compatible with Python 3.13+ and utilizes asynchronous operations.
A sample setup using memv:
from memv import Memory
from memv.embeddings import OpenAIEmbedAdapter
from memv.llm import PydanticAIAdapter
memory = Memory(
db_path="memory.db",
embedding_client=OpenAIEmbedAdapter(),
llm_client=PydanticAIAdapter("openai:gpt-4o-mini"),
)
async with memory:
await memory.add_exchange(
user_id="user-123",
user_message="I just started at Anthropic as a researcher.",
assistant_message="Congrats! What's your focus area?",
)
await memory.process("user-123")
result = await memory.retrieve("What does the user do?", user_id="user-123")
The project is currently at an early stage (v0.1.0), and feedback is encouraged, especially concerning the extraction approach and potential useful integrations.
📖 Read the full source: r/LocalLLaMA
👀 See Also

Handoffs Pattern in Claude Workflows: Two-File Split vs One-Doc Summary
Long Claude sessions break on context decay. Handoffs compress what matters and start fresh. Two approaches: Matt Pocock's single-doc handoff skill vs a two-file split with persistent narrative and ephemeral prompt.

VT Code: Open-Source Rust TUI Coding Agent with Multi-Provider Support and Agent Skills
VT Code is a Rust-based terminal UI (TUI) coding agent supporting Anthropic, OpenAI, Gemini, and Codex, with local inference via LM Studio and Ollama. It includes Agent Skills, Model Context Protocol, and Agent Client Protocol.

Alibaba's $10 monthly coding plan offers high-volume access to multiple AI models for OpenClaw users
For $10 per month, Alibaba's plan provides access to Qwen3.5-Plus, Kimi-K2.5, GLM-5, and MiniMax-M2.5 models with quotas of 1,200 requests per 5 hours, 9,000 per week, and 18,000 per month.

Tangent: Chrome Extension for Branching Claude Conversations
A free, open-source extension that lets you open side threads on Claude without losing your place.