DocMason: Local Agent Knowledge Base for Complex Office Files

What DocMason Does
DocMason is a local, file-based knowledge base system designed for deep research over private work documents. The core concept is "The repo is the app. Codex is the runtime." It compiles office files into structured evidence bundles that AI agents can reason over while maintaining strict provenance tracking.
Key Features from Source
- Handles multiple office document types: PPTX, DOCX, XLSX, PDFs, and even .EML files
- Extracts multimodal information including IT architecture diagrams and Excel sheet data
- Maintains document structure and visual semantics (slide layouts, presenter notes, spreadsheet references, formatting signals)
- Runs locally with no cloud ingestion or hidden backends
- Provides incremental knowledge base syncing when files are added or revised
- Enforces strict data contracts and provenance boundaries
How It Works
DocMason operates as a production-grade runtime that forces AI to respect original document structure. Instead of flattening complex files into unstructured text blobs, it creates deterministic file-based evidence and runs offline retrieval algorithms locally on your machine.
Getting Started
Two setup paths are described in the source:
Path A (Start Small):
- Drop work files into the
DocMason/original_doc/folder - Open the DocMason folder in Codex
- Ask questions naturally - DocMason guides through environment setup
- Approves prompts when building the knowledge base
Path B (Stage Entire Folders):
- Drop department-level folders into
DocMason/original_doc/ - Open in Codex and tell it: "Please prepare the DocMason environment."
- Then: "Please build the knowledge base."
- Once complete, ask complex research questions against the entire corpus
The system is designed so you don't need to memorize internal commands - just speak naturally to your AI agent within a valid workspace.
Technical Details
DocMason addresses specific limitations of existing document AI tools:
- Preserves visual layout, presenter notes, and chart-text relationships in slide decks
- Maintains multi-sheet references and nested tables in spreadsheets
- Retains formatting semantics like red text for "Risk" or indentation for hierarchies
- Enables cross-document reasoning for multi-part proposals
The repository structure includes adapters, knowledge_base, runtime, skills, and sample_corpus directories, with configuration managed through docmason.yaml and pyproject.toml files.
📖 Read the full source: HN AI Agents
👀 See Also

Madar: Local Context Compiler for Claude Code / Cursor — 78% Fewer Tokens on NestJS Repo
Madar is an open-source local context compiler for coding agents. On a NestJS + BullMQ repo (~800 files), it cut Claude Code input tokens by 78% and cost by 63% for an explanation task. Scoped graphs only.

SprintiQ: Open-Source Sprint Planning for Claude Code
SprintiQ is an open-source agile platform that acts as an orchestration layer for Claude Code, offering AI-powered user story generation, sprint planning, velocity tracking, and a CLI that syncs git activity to sprints in real time.

Quell Proxy Fixes Claude Code Scroll-Jumping on Windows
Quell is a Rust proxy that sits between your terminal and Claude Code, stripping clear-screen sequences that cause scroll position resets during long responses. It also adds Shift+Enter for newlines, security filtering, and full Unicode support.

Otterly: Route OpenClaw Through Your Claude Code Subscription
Otterly is a small npm package that exposes the local Claude CLI as an OpenAI-compatible HTTP server, letting you bill OpenClaw requests to your Claude Code subscription instead of pay-per-token API rates.