Memento v1.0: Local Persistent Memory for AI Coding Agents

What Memento v1.0 Does
Memento v1.0 provides a local-first persistent memory layer for AI coding agents. Everything runs on your machine — embeddings, storage, and search — with no cloud requirements or API keys needed after setup.
Key Technical Details
Embeddings: Uses all-MiniLM-L6-v2 via @xenova/transformers (384 dimensions) running fully offline. Optional cloud embeddings via environment variables for OpenAI (text-embedding-3-small) or Gemini (embedding-001).
Storage: Local JSON + HNSW index by default. Optional ChromaDB or Neo4j support.
Search: HNSW index for approximate nearest neighbor search (<50ms on 2000+ memories). Full BM25 implementation with k1=1.2, b=0.75 for keyword search. Hybrid mode combining 70% cosine similarity + 30% BM25.
Deduplication: SHA-256 + 0.92 cosine threshold.
Resilience features: Circuit breaker, write-ahead log, LRU cache.
Memory management: 347-day exponential decay on importance scores.
Setup and Usage
Install with: npx memento-memory setup
Migration tool: memory_migrate re-embeds your entire store when switching embedding providers — no data loss.
IDE Support and Tools
Multi-IDE compatibility: Claude Code, Cursor, Windsurf, OpenCode — all share the same local store.
17 MCP tools across save/recall/search/export/import/ingest/compact/graph/session lifecycle.
Privacy and Licensing
Zero telemetry — your architectural decisions and code patterns never leave your machine. Works without internet after setup. AGPL-3.0 licensed and self-hostable in one command.
📖 Read the full source: r/LocalLLaMA
👀 See Also

Replacing complex retrieval pipelines with simple git shell commands for LLM agents
A developer replaced their entire AI agent retrieval pipeline (sentence-transformers, rank-bm25, two-pass LLM pipeline) with a single tool that lets the agent execute read-only shell commands against a git repository, reducing Docker image size by ~3GB and eliminating timeout issues.

Claude-Skills Maintainer Seeks Feedback on 181 Agent Skills Library
Reza, maintainer of claude-skills, is asking the community for feedback on his open-source library containing 181 agent skills, 250 Python tools, and 15 agent personas that work across 11 AI coding tools. He's questioning whether the isolated skill approach is effective and wants input on missing skills, persona-based agents, and tool integrations.

Context-Engineered Study System for Claude Code Acts as Persistent Tutor
A developer built a study system using Claude Code that tracks progress across sessions, probes understanding, works through exercises, and adapts to learning styles. The system uses structured markdown files to shape agent behavior and includes tools for extracting textbook pages from PDFs.

Agent-Xray: Open-source tool for debugging AI agent failures from trace logs
Agent-Xray is an MIT-licensed open-source tool that analyzes AI agent trace logs to classify failures into categories like spin, tool_bug, and early_abort, and includes an enforcement mode to test fixes against adversarial challenges.