Distillery: A Claude Code Plugin for Persistent Team Context

Distillery is a plugin for Claude Code that creates shared, persistent context for development teams. It addresses the problem of teams losing knowledge when debugging sessions end or when decisions made months ago become difficult to find.
Core Commands
/distill— Capture decisions and reasoning mid-session. The whole team can search these later./recall— Find anything anyone on the team has captured using natural language search./pour— Synthesize coherent answers from scattered context across people and sessions. For example, asking "How does our auth system work?" pulls from multiple people's captured decisions and produces a narrative with citations.
Ambient Intelligence Features
The plugin includes ambient intelligence capabilities:
/watch— Point at GitHub repos, RSS feeds, or subreddits. It polls on a schedule and scores every item for relevance against your team's existing context using embedding similarity./radar— Provides a synthesized digest of what matters based on what the system learns your team cares about from captured context.
Team Deployment
Distillery uses a shared server with GitHub OAuth, allowing everyone to connect their Claude Code to the same knowledge base. Context captured by one person becomes searchable by everyone, creating compounding knowledge where each team member's captures improve everyone else's searches and syntheses.
Version 0.2.0 Updates
The latest release includes:
- Hybrid search (BM25 + vector with Reciprocal Rank Fusion)
- Auth audit logging
- UV support
The project is available on GitHub with a blog post explaining the development process.
📖 Read the full source: r/ClaudeAI
👀 See Also

Introducing NetViews 2.3: A Robust Network Diagnostic Tool for macOS
NetViews 2.3 combines host discovery, Wi-Fi insights, and real-time monitoring with a streamlined GUI for better network diagnostics on macOS.

mnemos: A Persistent Memory Layer for AI Coding Agents (Go, MCP-Native, No Python)
mnemos is a Go-based MCP-native memory layer for AI coding agents. The author built a verifier to measure lift: +40% aggregate on read-side scenarios, but only 53% write-side capture rate after iterative fixes.

Comparison of RunLobster vs Hosted OpenClaw Solutions
A developer tested RunLobster against KiwiClaw, xCloud, and self-hosted OpenClaw for 2 weeks each. RunLobster differs fundamentally as a product rather than just hosting, with 3,000 one-click integrations and memory that builds over time.

GLM 5 on Mac M3: Performance Observations for Agentic Coding
A user reports running GLM 5 via MLX 4-bit quantization on a Mac M3 with 512GB RAM, finding it usable for agentic coding with context under 50k tokens but noting significant slowdowns beyond that threshold.