Testing δ-Mem on Apple Silicon: MLX Implementation and Benchmarks

A Reddit user implemented the δ-mem research paper (arXiv 2605.12357) for Apple Silicon using mlx and OpenClaw integration. The paper improves model attention direction without context or LoRA, reporting 20% better answers in their tests. The implementation used Qwen3-4B-Instruct via mlx and custom adapters.
Benchmark Results (normalized mlx tests, Qwen3-4B-Instruct on MacMini 64GB):
- Synthetic paper-style: Plain 0.5129, δ-mem 0.5129 (1.00x)
- LoCoMo-10 mini: Plain 0.0500, δ-mem 0.1833 (3.67x)
- OpenClaw replay: Plain 0.5701, δ-mem 0.6667 (1.17x)
Latency costs (vs plain):
- Synthetic: 1.013x
- LoCoMo-10 mini: 1.33x query / 1.50x total
- OpenClaw replay: 1.30x
Key links:
- GitHub repo with adapter: delta-mem-mlx-sidecar-w-openclaw
- MLX adapter on Hugging Face: delta-mem-qwen3-4b-instruct-mlx-adapter
Takeaways:
- Synthetic probes were flat (1.00x), but LoCoMo-mini showed strong relative gains (3.67x).
- OpenClaw-style replay showed a practically meaningful improvement (6/8 → 7/8 probes passed, 1.17x).
- The user notes Apple Silicon cannot run CUDA efficiently, so results are lower than paper benchmarks. Paper benchmarks (Qwen3-4B-Instruct) showed avg 1.10x vs frozen backbone, MemoryAgentBench 1.31x, LoCoMo 1.20x.
- The user is seeking help (or funding ~$6k) to train an adapter for larger models like Qwen3.6:27B.
Who it's for: Developers running local LLM agents on Apple Silicon who want to experiment with δ-mem weight modulation to improve memory/context performance.
📖 Read the full source: r/LocalLLaMA
👀 See Also

Workflow orchestrator with AI CLI integration for sysadmin tasks
A developer built a file-based workflow orchestrator called 'workflow' that integrates with Claude Code, Codex CLI, and Gemini CLI. It generates, updates, fixes, and refines YAML workflows from natural language descriptions for sysadmin tasks.

Orloj: Declarative Orchestration Runtime for Multi-Agent AI Systems
Orloj v0.1.0 is an open-source orchestration runtime that lets you define AI agents, tools, policies, and workflows in YAML manifests with GitOps. It handles scheduling, execution, governance, and reliability for production multi-agent systems.

SmallClaw V1.0.3 Adds Webhooks, n8n Automation, and MCP Server Support
SmallClaw V1.0.3 introduces webhook endpoints for external service triggers, local automation workflows with n8n, and MCP server connections for tool integration. The update maintains the tool's focus on running with small local LLMs.

Monitor Your Claude AI Usage with a New Linux Taskbar Widget
A new Linux taskbar widget helps users track their Claude AI subscription usage in real-time, with color-coded feedback and easy installation.