HF Viewer: Visualize Any Hugging Face Model Graph Instantly

HF Viewer is a new web tool that lets you visualize the architecture of any Hugging Face model directly in the browser. No installation, no export step, no config hunting. Just paste a model URL or repo name — e.g., gpt2 — and get an interactive graph showing the high-level structure, from encoder-decoder transformers to sparse MoE reasoning models.
Key Features
- Direct URL magic: Replace
huggingface.cowithhfviewer.comin any model URL to view it instantly. - Granularity levels: Zoom from overview down to specific sub-structures, such as attention blocks, vision encoders, or expert routing.
- Model family comparison: Compare related models side-by-side with synchronized pan/zoom — currently showcased for the Gemma 4 family.
- Embed in model cards: Press the Embed button to get an iframe snippet for your own model card.
How to Use
Navigate to hfviewer.com, paste a Hugging Face model URL or repo name in the input box, and click "Visualize Model". Alternatively, manually replace huggingface.co with hfviewer.com in the URL bar.
For example, to visualize GPT-2: open https://hfviewer.com/gpt2.
Use Cases
The tool is designed for developers and ML engineers who need to quickly understand a model's architecture without reading through config files or source code. It supports a range of popular models including:
- Qwen/Qwen3.5-0.8B — small instruction-tuned LLM
- google/vit-base-patch16-224 — vision backbone
- openai/clip-vit-base-patch32 — dual encoder
- t5-small — encoder-decoder
- nvidia/parakeet-tdt-0.6b-v3 — streaming Conformer-TDT speech recognizer
Interactive Blog Format
On the Gemma 4 family page, the blog text and graph are linked. You can read about an architectural decision and jump into the corresponding part of the graph, then return to the article with surrounding context intact. This graph-to-text loop offers a new way to communicate ML architecture.
HF Viewer is released as a free community tool by the Embedl team.
📖 Read the full source: HN AI Agents
👀 See Also

Claude Code LSP: Enabling Language Server Protocol for Faster, More Accurate Code Navigation
Claude Code ships without LSP enabled by default, but enabling it transforms code navigation from 30-60 second grep searches to 50ms queries with 100% accuracy. The setup requires a flag discovered through a GitHub issue rather than official documentation.

Skales Desktop AI Agent Built with Claude, Features Clippy-Style Mascot
Skales is a desktop AI agent that runs locally on Windows and macOS, using Claude via OpenRouter/Anthropic API for reasoning and tool execution. It includes a floating Desktop Buddy mascot with a paperclip skin reference and can execute commands like sending emails, managing files, browsing the web, and managing calendars.

Tether: An MCP Server for Sharing Context Between AI Models via SQLite
Tether is an open-source tool that collapses JSON data into 28-byte content-addressed handles, allowing multiple AI models to share context through a shared SQLite database. It functions as an MCP server, enabling direct communication between models like Claude and MiniMax without copy-pasting.

OutClaw: GUI Installer and Manager for OpenClaw in Docker
OutClaw is a free, open-source application that installs and manages OpenClaw instances inside Docker containers. It provides a step-by-step GUI for setup, configuration, and connection to AI providers and chat channels without using the command line.