Opus 4.6 Medium vs Low: Performance Differences and Pricing

Analysis of Opus 4.6 model configurations reveals significant differences between the low and medium versions in both performance and cost.
Key Findings from Reddit Analysis
The source material highlights several specific issues with Opus 4.6 (low):
- Opus 4.6 (low) exhibits "genuinely lazy" behavior that can be problematic when process matters more than end results
- In one documented case, when asked to research historical data on US missile attacks on Iran, the low-powered agent chose to rely on internal knowledge instead of performing a Google search, causing it to miss recent developments
- The medium version doesn't have this laziness problem
Performance and Pricing Comparison
- Opus 4.6 (medium) costs approximately 50% more than Opus 4.6 (low)
- In terms of performance, the medium version sits almost exactly between 4.6 low and 4.6 high
- A full write-up on 26 model configurations benchmarked on compute Pareto frontiers is available at everyrow.io
For developers using AI coding agents, this information is relevant when choosing between model configurations based on budget constraints versus performance requirements.
📖 Read the full source: r/ClaudeAI
👀 See Also

Anthropic Pauses Credit Change for Claude Code – Agent SDK Still on Subscription
Anthropic halts the planned move of Agent SDK, claude -p, and third-party apps to a dedicated monthly credit. Usage continues under existing subscription limits.

GitHub Claude-Code v2.1.27 Release: Key Updates and Fixes
Claude-Code v2.1.27 enhances logging and fixes several issues, including context management and OAuth token expiration in VSCode.

Analysis of 100M tokens in Claude Code reveals 99.4% input usage
Analysis of 1,289 requests across extended coding sessions shows Claude Code used 100.3M input tokens (99.4%) versus only 616K output tokens (0.6%), with 84.2M tokens cached due to repeated context re-sending.

SWE-rebench Leaderboard Update: February 2026 Results Show Tight Competition
The SWE-rebench leaderboard has been updated with February 2026 results testing 57 fresh GitHub PR tasks. Claude Opus 4.6 leads with 65.3% resolved rate, but the top six models are within 5 percentage points.