engram v3.4.0 Adds Anthropic Plugin to Keep Claude Code Running Under New Rate Limits

✍️ OpenClawRadar📅 Published: May 18, 2026🔗 Source
engram v3.4.0 Adds Anthropic Plugin to Keep Claude Code Running Under New Rate Limits
Ad

engram v3.4.0 addresses the recent rate-limit reductions and the impending removal of Claude Code from the Pro tier by exposing a dedicated Anthropic plugin. The plugin bundles an MCP server config that instantiates a shared memory layer locally, surviving file edits and IDE switches without extra latency.

Key Features

  • Three new skills accessible via slash commands in Claude Code: /engram:cost for token spend queries, /engram:query for fast context retrieval, and /engram:mistakes to surface recent execution errors.
  • Zero-config MCP integration — the MCP server runs locally, so the context spine is instantiated the first time a skill runs, with no additional setup.
  • Cross-IDE persistence — the shared memory layer persists across file edits and even across different IDEs, enabling continuity.

Installation

CLI (one line):

npm install -g engramx@latest
engram setup # detects Claude Code automatically

Via Claude Code marketplace: Once the listing appears, run /plugin install engram.

Ad

What It Solves

Claude Code users have faced sudden rate-limit reductions with the product's looming removal from the Pro tier. engram's plugin provides a local, latency-free memory layer that helps manage API consumption (via cost queries) and recover from errors quickly (via mistake surfacing). The MCP server runs locally, so no external dependencies are introduced.

Who It's For

Developers who rely on Claude Code and need to work around tighter rate limits while maintaining continuity across sessions.

Resources

📖 Read the full source: r/ClaudeAI

Ad

👀 See Also

Custom llama.cpp Backend Offloads LLM Matrix Multiplication to AMD XDNA2 NPU on Ryzen AI MAX 385
Tools

Custom llama.cpp Backend Offloads LLM Matrix Multiplication to AMD XDNA2 NPU on Ryzen AI MAX 385

A developer built a custom llama.cpp backend that dispatches GEMM operations directly to the AMD XDNA2 NPU on Ryzen AI MAX 385 (Strix Halo), achieving 43.7 t/s decode at 0.947 J/tok with Meta-Llama-3.1-8B-Instruct Q4_K_M. The NPU decode path saves ~10W versus Vulkan-only while matching decode throughput.

OpenClawRadar
Open Source Agent Skill for TypeScript, React, and Next.js Patterns
Tools

Open Source Agent Skill for TypeScript, React, and Next.js Patterns

A developer has released a 4,000-line, 17-file structured markdown reference designed for AI agents like Claude Code to follow when generating or reviewing TypeScript, React, and Next.js code. It addresses common issues like improper API response validation and misuse of 'use client' directives.

OpenClawRadar
Claude Code Ultracode Mode Spawns 70-Agent Pipeline for Deep Search
Tools

Claude Code Ultracode Mode Spawns 70-Agent Pipeline for Deep Search

A single 'deep search' request in Claude Code's ultracode mode auto-generated a 4-phase pipeline with ~70 agents, each fetching and cross-checking projects independently. The orchestrator script keeps intermediate results out of the context window, preventing context overload.

OpenClawRadar
GSD-Lite: A State Machine for Claude Code That Enforces TDD and Prevents Test Skipping
Tools

GSD-Lite: A State Machine for Claude Code That Enforces TDD and Prevents Test Skipping

GSD-Lite is an open-source MCP server that adds a 12-state workflow machine to Claude Code, enforcing test-driven development with specific anti-rationalization prompts and separate agent contexts for execution, review, and debugging.

OpenClawRadar