Critique of MCP's Abstraction Boundary and Service Integration Approach

A Reddit discussion on r/ClaudeAI examines where MCP draws its abstraction boundary and argues it gets this wrong in a way that goes beyond typical implementation criticisms like security, token bloat, or transport issues.
Core Argument About Service Integration
The post identifies three separate concerns when an agent needs to work with a service: API access, efficient tooling that wraps it, and domain knowledge about how to use it well. According to the source, MCP bundles all three into one layer, resulting in a limited subset of what the underlying API can actually do.
Lattice as a Concrete Example
The discussion uses Lattice as a specific example. Their web client is powered by a full GraphQL API that covers everything an employee would want to do. However, their public API only covers HR admin workflows. The post argues that MCP incentivizes services to build yet another limited interface rather than just opening up the APIs they already have.
Proposed Alternative Approach
The author suggests a better path: services making their primary APIs universally accessible, noting that the authentication problem is already solved by OAuth 2.0 with PKCE. Domain knowledge should be distributed as agent skills rather than baked into MCP tool definitions.
The full post is available at tomyandell.dev/blog/my-problem-with-mcp, and the Reddit discussion invites others to share their thoughts on whether this critique of the abstraction boundary is correct.
📖 Read the full source: r/ClaudeAI
👀 See Also

Hy3 LLM Tops OpenRouter Rankings: Cheapest Model or Something Else?
Hy3 preview, a Tencent open-source LLM, surged to the top of OpenRouter's model rankings by token usage, surpassing Claude and DeepSeek V4 Flash. Priced at $0.066/1M input tokens, it's the cheapest major model, but benchmarks show quality far below leaders.

Autoresearch Pushes Qwen3.5-397B to 20.34 tok/s on M5 Max via SSD Streaming
A developer achieved 20.34 tokens/second inference speed for the 209GB Qwen3.5-397B model on a MacBook Pro M5 Max with 128GB RAM using SSD streaming and 36 systematic experiments. The result represents a 2x speedup over the M5 Max baseline and 4.67x over the original M3 Max result.

Google DeepMind Workers Vote to Unionize Over Military AI Deals
London-based Google DeepMind employees voted to unionize, demanding Google halt AI contracts with US and Israeli militaries, citing concerns over ethical guidelines removal.

Startups Report Spending More on AI Compute Than Human Salaries
AI startups like Swan AI report monthly AI compute bills exceeding $113k, with CEOs describing this as 'tokenmaxxing' where AI spending replaces traditional headcount budgets.