Librarian MCP: Local AI Server for Persistent Context with Documents

What Librarian MCP Does
Librarian MCP is an open-source Model Context Protocol server that plugs into Jan, LM Studio, or Claude Desktop, turning your local chat window into an interactive research assistant. It solves the problem of document collections that are too large for context windows but too private to send to cloud APIs.
Key Features
- Runs 100% locally with Qwen, GLM, Llama, or any local model
- Remembers everything across your entire conversation (persistent context)
- Searches semantically (finds concepts, not just keywords)
- Writes analysis reports to a sandboxed workspace (you review before applying)
- Works on ANY document collection - code repos, research papers, medical records, legal contracts, Obsidian vaults
- Adopts specialist personas - debugging analyst, compliance expert, legal analyst, knowledge synthesizer
Quick Start Installation
Three-step setup:
git clone https://github.com/orangelightening/Librarian.git && cd Librarian && ./install.shCopy the config output to Jan's MCP settings, then open a new chat.
How It Works
Point it at your documents (any format), open Jan/LM Studio/Claude Desktop, and start chatting with your library. The Librarian maintains context across your entire conversation, building increasingly sophisticated understanding as you chat.
Privacy and Security
- No API calls required
- No data leaves your machine
- Write access is sandboxed to /librarian/ only (can't modify your actual documents)
- Described as having 7 security layers
Technical Details
- Chonkie backend (intelligent semantic chunking)
- ChromaDB vector storage
- 14 production tools (search, sync, read, write, execute, etc.)
- Works with: Jan, LM Studio, Claude Desktop, any MCP client
Real-World Use Cases
- Debugging: "Trace why document sync is failing" → Root cause with code paths
- Legal: "Find inconsistent contract clauses" → Risk assessment report
- Medical: "Validate policies against HIPAA" → Compliance audit
- Obsidian: "Find connections across my notes" → Knowledge map
Perfect for: medical records, legal contracts, corporate data, personal knowledge bases.
📖 Read the full source: r/LocalLLaMA
👀 See Also

Agent Kernel: Three Markdown Files for Stateful AI Agents
Agent Kernel provides three markdown files that enable stateful behavior in AI coding agents without databases or custom frameworks. It works with OpenCode, Claude Code, Codex, Cursor, Windsurf, and similar tools.

MCP Server for Italian Train Data: Real-Time Delays, Departures, and Schedules in Claude
A developer built an unofficial MCP server for Trenitalia that provides five tools for querying Italian train data through Claude, including real-time departure/arrival boards, train tracking, and schedules with live delay enrichment.

Claude for Design Work: How to Stop Repeating the Same Taste Arguments Every Session
A developer running client work through Claude describes the core problem: Claude has no memory of rejected design decisions, leading to generic outputs and inconsistent brand identity.

VTCode: A Rust TUI Coding Agent That Aggressively Trims Context with AST-Level Chunking
VTCode is an open-source Rust TUI coding agent that aggressively trims context using AST-level chunking via ripgrep and ast-grep. It supports custom OpenAI-compatible providers, sandboxing with macOS Seatbelt and Linux Landlock, and tree-sitter-bash validation on generated commands.