LLM Matrix: Community-Voted Model Comparisons Built with Claude Code

A developer has created LLM Matrix, a website that allows users to browse and vote on large language models across multiple dimensions. The tool addresses concerns about centralized benchmark sites by implementing community-driven rankings.
What LLM Matrix Does
- Browse LLM scores across 2 to N dimensions simultaneously
- Users vote on models, and those votes shape the rankings
- Initial data seeded with only 20 votes per model based on aggregated scores from public internet sources
- Remaining votes and rankings determined by community input
Development Details
The entire project was built using Claude Code. The developer specifically mentioned two plugins that were essential to the development:
- production-grade plugin:
https://github.com/nagisanzenin/claude-code-production-grade-plugin - claude-mem plugin:
https://github.com/thedotmack/claude-mem
The site is currently hosted at llm-matrix.vercel.app and represents an alternative approach to LLM evaluation that prioritizes community consensus over potentially biased centralized metrics.
📖 Read the full source: r/ClaudeAI
👀 See Also

WinRemote MCP: Open Source MCP Server for Full Control of Windows Desktops
WinRemote MCP provides AI agents with full control over Windows desktops, allowing for UI detection, file operations, registry access, and more, utilizing over 40 tools.

Natural Language Autoencoders: Turning Claude's Internal Representations into Text
Transformer Circuits Thread publishes Natural Language Autoencoders that decode Claude's internal activations into readable text. GitHub repo and interactive demo available.

IUM: MCP Symbol Indexer Cuts AI Agent Token Usage by 15.9x vs grep
IUM indexes codebases into an SQLite matrix of symbol events, exposing exact file:line coordinates, call graph tracing, and semantic search via MCP. Benchmarked against DataFusion (1,538 files) showing 15.9x fewer tokens than grep for equivalent queries.

Multi-Model Council Workflow for AI Coding Agents
A developer built a web tool that runs coding tasks through three AI models—GPT-4o as architect, Claude as skeptic, and Gemini as synthesizer—before passing them to coding agents. The tool generates a PLAN.md with explicit constraints and requires users to bring their own API keys.