ScreenMind: Local-First AI Memory That Indexes Your Entire Computer Activity

ScreenMind is a local-first AI memory system that continuously captures your screen, transcribes meetings, and indexes voice notes, building a persistent, searchable timeline of everything you do on your computer. It uses perceptual hashing to only trigger when content changes, then runs each frame through Gemma 4 E2B via llama.cpp for vision analysis, chat, and audio processing.
Key Features
- Screen capture with perceptual hashing — only stores frames when content actually changes
- Searchable timeline — query past activity: "that error message from earlier," "what was I working on at 3pm?"
- Chat with your history — persistent AI context from your entire session
- Meeting transcription — auto-detects Zoom, Teams, and Google Meet
- Voice memos — processed via Gemma 4's audio encoder
- Natural language automations — write them in plain English Markdown
- MCP integration — connect to Claude and Cursor
Technical Stack
- Models: Gemma 4 E2B (handles vision, chat, audio)
- Backend: Python + FastAPI
- Storage: SQLite
- Inference: llama.cpp with Q4 quantization
- Hardware: 4GB+ VRAM
The author notes that GPU scheduling between vision, chat, and audio tasks is the main inference optimization challenge. The project is still workflow-driven rather than fully autonomous — retrieval quality and onboarding friction are areas needing improvement.
GitHub: ayushh0110/ScreenMind
📖 Read the full source: r/LocalLLaMA
👀 See Also

Phantom: A Persistent AI Agent Built with Claude's Agent SDK
Phantom is an open-source Bun/TypeScript process that wraps Claude's Agent SDK (Opus 4.6) with persistent vector memory, a self-evolution engine, and an MCP server interface. It runs continuously on its own VM or Docker Compose and communicates via Slack.

OpenClaw Agent Gains Phone Call Capability Through Custom Skill
A developer created a custom skill for self-hosted OpenClaw agents that enables phone call functionality, allowing the agent to initiate calls based on triggers like build completions or server outages. The implementation provides voice interaction with full chat capabilities including web searches and alert setup.

ClawCode: Cleanroom Rust Rewrite of Leaked Claude Code
ClawCode is a cleanroom rewrite of the leaked Claude Code source code, implemented in Rust. The project emerged following Anthropic's Claude Code leak and is being compared to OpenCode for end-to-end task performance.

Bifrost LLM Gateway: 11 Microsecond Overhead, Single Binary in Go
Bifrost is an open-source LLM proxy written in Go that routes requests to OpenAI, Anthropic, Azure, and Bedrock with 11 microsecond overhead per request, handling 5,000 RPS on a $20/month VPS.