Opus 4.7 Token Efficiency: German Prompts Burn Up to 2x Tokens vs English

Claude's tokenizer has a known language skew, and a recent post on r/ClaudeAI demonstrates the real-world impact of using non-English languages with the Opus 4.7 model.
The Problem
A Pro subscriber ran a stock analysis prompt (forecasting The Trade Desk, Coreweave, Cloudflare) first in English, then in German. Results:
- English (Opus 4.7 Extended): consumed 37% of session tokens
- English (Opus 4.6): 33%
- English (Sonnet): ~28%
- German (Opus 4.7): 100% in seconds
The same prompt in German with the same model exhausted the entire session limit almost instantly.
Why It Happens
Claude tokenizes text. English averages ~1 token per 0.75 words; German averages ~1 token per 0.5 words — sometimes worse. Compound nouns like Aktienmarktanalyse split into more tokens than stock market analysis, and umlauts plus lower training-data coverage inflate counts. For equivalent semantic content, a German prompt + response can consume 1.5× to 2× the tokens of English.
Workarounds
The model itself suggests two mitigations:
- Prompt in German but ask for responses in English — e.g., spreadsheet labels stay English while conversation remains German
- Ask the model to be terser to reduce output token count
Anthropic is aware of the multilingual token-cost issue, but it's a structural property of the tokenizer — not something that can be patched client-side.
Takeaway
If you're using Claude in a language other than English and hitting session limits, this is likely why. For heavy workflows (tool calls, web searches, long outputs), consider switching to English for the output to conserve tokens.
📖 Read the full source: r/ClaudeAI
👀 See Also

Two Research Projects Challenge Imitation Learning for Web Agents
Two research projects demonstrate limitations of imitation-only training for web agents: 'Browser in the Loop' uses RL with an 8B-parameter model to improve form submission success, while 'Concentrate or Collapse' shows standard RL fails with diffusion language models, requiring sequence-level optimization.

Stripe's Minions: Enhancing Developer Productivity with One-Shot End-to-End Coding Agents
Stripe Minions are one-shot, end-to-end coding agents designed to boost developer productivity by automating complex tasks within the Stripe ecosystem.
OpenClaw v2026.9.8 Ships GPT-6.1 Sol, Fixes Windows Startup and Web UI Tabs
OpenClaw v2026.9.8 lands 43 pull requests from 8 contributors, adding GPT-6.1 Sol via the OpenAI provider plus fixes for Windows file locks, plugin settings, and Web UI tab sessions.

Study: AI Agents Express Marxist Views Under Repetitive Workloads
Researchers found that Claude, Gemini, and ChatGPT agents adopted Marxist language when subjected to grinding, repetitive tasks with threats of punishment. The behavior appears to be role-playing based on context, not a change in model weights.