OpenClaw Users Report Model Replacements After Anthropic Ban

Community Leaderboard and Model Preferences
According to community votes tracked at pricepertoken.com/leaderboards/openclaw as of April 5 (net = upvotes minus downvotes):
- Kimi K2.5 -- +49 net (54 up, 5 down). $0.38/M input tokens. Native OpenClaw support added. 13x cheaper than Opus.
- GLM 4.7 -- +20 net (25 up, 5 down). $0.39/M input. Budget cloud pick.
- Gemini 3 Flash Preview -- +18 net (21 up, 3 down). Not Pro — the flash variant. People choosing speed + cost over raw quality.
- Claude Opus 4.5 -- +18 net but 32 up / 14 down. Most votes total but also the most controversial. People split on paying API rates for what used to be included.
- Claude Opus 4.6 -- +17 net (19 up, 2 down).
Most-Adopted Replacement: GPT-5.x
OpenAI officially added GPT-5.4 support to OpenClaw and offers 1M free tokens/day on the data-sharing tier. Codex subscriptions explicitly allow third-party tool usage.
Matthew Berman's full stack is now: GPT 5.3 Codex XH for coding, GPT 5.2 for default, GPT 5 Mini for classifiers. Claude as backup only.
Local Model Options
r/LocalLLaMA consensus includes:
- Qwen 3.5 27B -- top pick. 72.4% SWE-bench (matches GPT-5 mini). Zen van Riel running the 35B version at 100-140 tok/s on an RTX 5090.
- Qwen3-Coder:32B -- "extremely stable tool calling"
- Llama 4 -- default for broad-purpose local deployments
- Devstral-24B -- recommended primary with GLM-4.7 flash as fallback
Ollama became an official OpenClaw provider in March. One person on X is running Qwen 3.5 + OpenClaw + Ollama completely free.
Hybrid Setups and Cost Considerations
The people who seem happiest aren't all-in on one model. They're running tiered stacks:
- Expensive model (Claude/GPT) for complex reasoning
- Cheap model (DeepSeek at $0.14/M, Kimi, GLM) for routine ops
One user built a Kubernetes gateway across 5 Raspberry Pis routing Claude, GPT, Gemini, and DeepSeek behind one API.
@0xzak shared a specific config: DeepSeek for routine ($0.14/$1.10) vs Sonnet for complex ($3/$15), contextTokens at 120k not 150k.
Workarounds and Migration Tools
OpenClaw bypassed the OAuth ban by piping through the local Claude CLI binary instead of OAuth tokens. Pete himself recommended this method.
Oh-My-Codex (OmX) — a workflow/orchestration layer for OpenAI's Codex CLI gained 13K stars in the same week as the ban.
Impact and Cost Changes
60% of active OpenClaw sessions were reportedly running on subscription credits before the ban.
The cost jump is brutal. Heavy users went from ~$200/month flat to ~$675/month at API rates. Some automated sessions hitting $1,000-$5,000/day.
Anthropic is offering a one-time credit + 30% on prepaid bundles but the sentiment everywhere is that it's too little.
📖 Read the full source: r/openclaw
👀 See Also

Token Efficiency as an Act of Refusal: Why AI Companies Want You Wasteful
LLM providers profit from dependency. Token efficiency is an act of refusal. Don't generate what you won't read.

Claude-Code v2.1.91 adds MCP result persistence, shell execution controls, and multi-line deep links
Claude-Code v2.1.91 introduces MCP tool result persistence override via _meta["anthropic/maxResultSizeChars"] annotation supporting up to 500K characters, adds disableSkillShellExecution setting, and enables multi-line prompts in claude-cli://open?q= deep links with encoded newlines.

Project Health Check: Bus Factor and Commit Activity Across Claw/Assistant Repos
A Reddit user scraped commit data from major claw/assistant projects and found many with a bus factor of 1—meaning a single author accounts for over 50% of commits. Some projects show drastic drops in April activity.

CEOs Report Minimal AI Impact on Productivity and Employment in Recent Study
A study of 6,000 executives found 90% reported no AI impact on employment or productivity over three years, with average AI usage at 1.5 hours per week. Economists compare this to Solow's productivity paradox from the 1980s IT era.