Black LLAB: Open-Source Architecture for Dynamic Model Routing and Docker-Sandboxed AI Agents

A developer has released Black LLAB, an open-source project that attempts to replicate frontier AI lab systems for autonomous task execution. The system addresses two main problems: manually deciding which model to use for different prompts and safely executing AI agent code.
Architecture Components
The system consists of several key components:
- Dynamic Complexity Routing: Uses Mistral 3B Instruct to grade prompts on a scale of 1-100. Simple questions get routed to fast/cheap models; complex coding tasks get routed to heavy models with "Lost in the Middle" XML context shaping.
- Docker-Sandboxed Agents: Integrates OpenClaw to deploy agents in dedicated, isolated Docker containers. Agents can write files, scrape the web, and execute code without touching the host OS.
- Advanced Hybrid RAG: Builds a persistent Knowledge Graph using NetworkX and uses a Cross-Encoder for precise context retrieval beyond standard vector search.
- Live Web & Vision: Integrates with local SearxNG for web scraping and Pix2Text for local vision/OCR.
- Budget Guardrails: Includes a daily spend limit slider to prevent cloud API overages.
Model Lineup
The system uses multiple models for different purposes:
- Routing/Logic: Mistral 3B & Qwen 3.5 9B (Local)
- Midrange/Speed: Xiaomi MiMo Flash
- Heavy Lifting (Failover): Claude Opus & Perplexity Sonar
Tech Stack
The project is built with FastAPI, Python, NetworkX, ChromaDB, Docker, Ollama, Playwright, and a vanilla HTML/JS terminal-inspired UI.
The developer describes themselves as "more a mechanical engineer than software" and is seeking senior developer feedback on the architecture, particularly the Docker sandboxing approach. The project is available on GitHub for independent researchers who want to run autonomous tasks without being locked to a single provider.
📖 Read the full source: r/openclaw
👀 See Also

2026 Hermes Agent Alternatives Roundup: Self-Hosted Options from OpenClaw to memU Bot
A developer who has been running Hermes since launch tested every self-hosted and managed alternative after the ClawHub security mess. Key findings: OpenClaw (370k stars) but 9 CVEs in 4 days and ~20% malicious packages; TrustClaw rebuilt with OAuth/sandboxing; nanobot at ~4K lines Python with MCP; memU Bot with unique structured memory. Managed options include Perplexity Computer (19 models, $200/mo), Claude Cowork (opens real Mac apps), and KimiClaw (40GB RAG, locked to K2.5, Chinese data law). Full roundup at source.

Inline Visualizer: Local AI Models Can Now Render Interactive HTML Visualizations
Inline Visualizer is a BSD-3 licensed plugin for Open WebUI that enables any local AI model with tool calling support to render interactive HTML/SVG visualizations directly in chat, with a JavaScript bridge allowing elements to send messages back to the AI.

Persistent Memory for Claude: Local Stack with MCP, 39ms Retrieval, 82% Token Reduction
A developer built a persistent memory layer for Claude using local vector search (Qdrant + Qwen3) and MCP integration, achieving 82% token reduction, 39ms hot-path retrieval, and session crystallization via L4 nodes.

Session Siphon: Open Source Tool Consolidates AI Coding Agent Conversations
Session Siphon is a free, open source tool that consolidates and indexes conversation history from multiple AI coding agents across different providers and machines. The developer created it using Claude to solve the problem of tracking conversations across different platforms.