Model Routing Baselines for Claude and OpenAI Usage

A developer on r/openclaw shared their current model routing baselines for working with Claude and OpenAI models. The setup assigns specific models to different task types based on complexity and cost considerations.
Primary Model Assignments
The routing table includes:
- Default/General tasks: Claude Haiku 4.5 - "Fast, cheap baseline for most tasks"
- Architecture & Design: Claude Sonnet 4.6 - "Complex reasoning needed"
- Security Analysis: Claude Sonnet 4.6 - "Injection resistance matters"
- Debugging: Claude Sonnet 4.6 - "After 2 failed Haiku attempts"
- Major Decisions: Claude Sonnet 4.6 - "Multi-project impact"
- All Coding Tasks: ChatGPT 5.3 Codex - "Writing, debugging, review, codebase architecture"
- Advanced Reasoning: Claude Opus 4.6 - "Only if Sonnet can't solve it"
Fallback Strategy
When Anthropic models are unavailable:
- Standard tasks → GPT-5 Mini
- Complex tasks → GPT-5.4
Cost Optimization Rules
The developer specifies never using premium models for:
- File reads/writes
- Simple questions
- Status updates
- Formatting
- Anything Haiku handles in one shot
📖 Read the full source: r/openclaw
👀 See Also

How to safely run llama.cpp native tools (exec_shell_command) with multi-sandboxing on Linux
A practical guide to enabling llama.cpp native tools, especially exec_shell_command, and running them inside multiple sandboxes (Firejail + tiny Alpine VM) for safe web fetching and command execution via the llama-server web UI.

Qwen 3.5 122B MoE at 35 t/s on a Single 3090 with ik_llama.cpp MTP
A local stack running Qwen 3.5 122B MoE on a single 3090 at 35 t/s using ik_llama.cpp's fused MoE ops for MTP. Stock llama.cpp showed only +4% improvement; ik's fork yields +20%.
Automating OpenClaw Upgrades with an AI Agent: A Field-Tested Playbook
A developer shares how they trained a Hermes AI agent to handle OpenClaw upgrades, cutting a multi-hour, white-knuckle migration down to 'run the playbook, hit two known snags.'

Cost-Effective OpenClaw Multi-Agent Setup Using Subscription Models
A Reddit user describes routing all OpenClaw multi-agent operations through existing $200 Anthropic Pro Max and $200 ChatGPT OpenAI Codex subscriptions instead of raw API calls, using cheaper Anthropic models for simple agents and more complex models for others.