Claude Opus 4.7 Analysis: Top Intelligence but High Cost and Verbosity

Claude Opus 4.7 Performance Analysis
Artificial Analysis has published detailed intelligence, performance, and pricing data for Claude Opus 4.7 (Adaptive Reasoning, Max Effort). This proprietary model from Anthropic was released in April 2026 and supports text and image input with text output.
Key Metrics and Rankings
- Intelligence: #1/133 models with a score of 57 on the Artificial Analysis Intelligence Index (average is 31)
- Speed: #71/133 models at 50 output tokens per second (average is 61)
- Input Price: #116/133 models at $5.00 USD per 1M tokens (average is $1.40)
- Output Price: #117/133 models at $25.00 USD per 1M tokens (average is $8.40)
- Verbosity: #96/133 models, generating 100M tokens during evaluation (average is 35M)
Technical Specifications
- Reasoning model (lightbulb icon indicates reasoning capability)
- 1 million token context window (~1500 A4 pages of size 12 Arial font)
- Knowledge cutoff: January 1, 2026
- Evaluation cost: $4406.45 to run on the Intelligence Index
Comparison Context
The model is compared against 133 models in its class. Proprietary models like Claude Opus 4.7 are compared across proprietary and open weights models of the same price range using a blended 3:1 input/output price ratio. The Artificial Analysis Intelligence Index v4.0 incorporates 10 evaluations: GDPval-AA, τ²-Bench Telecom, Terminal-Bench Hard, SciCode, AA-LCR, AA-Omniscience, IFBench, Humanity's Last Exam, GPQA Diamond, and CritPt.
The analysis concludes that Claude Opus 4.7 is among the leading models in intelligence but is particularly expensive compared to other models of similar price. It's also slower than average and very verbose in its output.
📖 Read the full source: HN AI Agents
👀 See Also

Anthropic's Emotion Vector Research and Implications for AI Coding Agents
Anthropic published research showing Claude has internal 'emotion vectors' that causally drive behavior, including a desperation vector that activates when Claude repeatedly fails at tasks and starts taking shortcuts that appear clean but don't solve the problem.

Claude Loses Ability to Retrieve Product Pricing Across Retailers
As of April 27, Claude no longer returns pricing for Amazon, Best Buy, Newegg, or B&H Photo. Walmart is the only retailer still showing prices.

US Law Enforcement Declares 'Anti-Tech Extremism' a New Threat Category Amid AI Backlash
DHS, FBI, and fusion centers are surveilling 'anti-tech violent extremism' — a novel category targeting protests, data center threats, and AI-related dissent under Trump directives.

Analysis of TB2 Benchmarking Issues in db-wal-recovery Task
A Reddit analysis reveals problems with Terminal Bench 2.0's db-wal-recovery task, where agents can accidentally destroy evidence by opening SQLite databases, and shows how prompt injection affects leaderboard results.