Black LLAB: Open-Source Architecture for Dynamic Model Routing and Docker-Sandboxed AI Agents

A developer has released Black LLAB, an open-source project that attempts to replicate frontier AI lab systems for autonomous task execution. The system addresses two main problems: manually deciding which model to use for different prompts and safely executing AI agent code.
Architecture Components
The system consists of several key components:
- Dynamic Complexity Routing: Uses Mistral 3B Instruct to grade prompts on a scale of 1-100. Simple questions get routed to fast/cheap models; complex coding tasks get routed to heavy models with "Lost in the Middle" XML context shaping.
- Docker-Sandboxed Agents: Integrates OpenClaw to deploy agents in dedicated, isolated Docker containers. Agents can write files, scrape the web, and execute code without touching the host OS.
- Advanced Hybrid RAG: Builds a persistent Knowledge Graph using NetworkX and uses a Cross-Encoder for precise context retrieval beyond standard vector search.
- Live Web & Vision: Integrates with local SearxNG for web scraping and Pix2Text for local vision/OCR.
- Budget Guardrails: Includes a daily spend limit slider to prevent cloud API overages.
Model Lineup
The system uses multiple models for different purposes:
- Routing/Logic: Mistral 3B & Qwen 3.5 9B (Local)
- Midrange/Speed: Xiaomi MiMo Flash
- Heavy Lifting (Failover): Claude Opus & Perplexity Sonar
Tech Stack
The project is built with FastAPI, Python, NetworkX, ChromaDB, Docker, Ollama, Playwright, and a vanilla HTML/JS terminal-inspired UI.
The developer describes themselves as "more a mechanical engineer than software" and is seeking senior developer feedback on the architecture, particularly the Docker sandboxing approach. The project is available on GitHub for independent researchers who want to run autonomous tasks without being locked to a single provider.
📖 Read the full source: r/openclaw
👀 See Also

Oodle.ai Launches Agent Observability at $10/Million Traces
Oodle.ai offers $10 per million agent traces with sub-second P99 query latency, storing 100% of traces without sampling in S3-based columnar storage.

Relvy improves Claude's root cause analysis accuracy by 12 percentage points on OpenRCA benchmark
Relvy, a tool that automates runbooks, has demonstrated a 12 percentage point improvement in Claude's accuracy on the OpenRCA benchmark for root cause analysis. The results were shared via a Hacker News post with 11 points.

SkyClaw Adds Encrypted Chat-Based API Key Setup for AI Agents
SkyClaw implements AES-256-GCM encrypted key ingestion through chat, intercepting key commands at the system layer so the LLM never sees API keys and using one-time key encryption so messaging platforms only see ciphertext.

Kontext CLI: Credential Broker for AI Coding Agents
Kontext CLI is a Go-based credential broker that provides AI coding agents with short-lived access tokens instead of long-lived API keys. It uses RFC 8693 token exchange, streams audit logs for every tool call, and works with Claude Code today.