Balloon That Pops When Claude Finishes: Physical Agent UI with whisper.cpp
A developer built a desktop balloon UI that inflates as you speak a task, then floats across your screen while a real Claude Code agent runs in the background. When the agent finishes, the balloon pops with confetti and a result card. If it fails, the balloon deflates sadly.
How It Works
- Voice input: Hold a nozzle on a helium tank icon, speak a task like "Clean up my Downloads folder and sort everything into subfolders" or "make me a todo list app." The balloon inflates with your voice, and the transcript appears directly on it.
- Agent execution: When you let go, the balloon floats across the screen while a real Claude Code agent runs the task in the background. The balloon fidgets whenever the agent is actively doing something.
- Live console: Double-click the balloon to drop down a tiny console showing every thought and tool call live.
- Parallel agents: Speak multiple tasks and have several balloons floating at once, each running its own agent in parallel.
- Drag and park: Drag a balloon by its tag to park it somewhere.
Technical Details
- Transcription: Runs locally using
whisper.cpp— no cloud dependency. - Agent backend: The agents run through the Claude Agent SDK using your existing Claude Code login, so there's no setup and no separate bill.
- Licensing: Completely free and open source.
Why This Exists
The creator notes that agents still feel intimidating to most people, and "nobody is scared of a balloon." This project is the opposite end of the same idea as a physical gear shifter for switching models — making agent software feel physical.
📖 Read the full source: r/ClaudeAI
👀 See Also

context-os: Open-source tool reduces Claude Code token consumption by 27-42%
context-os is a local context optimizer that hooks into Claude Code automatically, compressing tool output before Claude sees it and reducing token consumption by 27-42% depending on content type.

Building a Self-Updating Writing Style Guide for AI-Assisted Content
A team building a voice extraction platform called Noren has developed a 117-line Markdown style guide that rewrites itself after every published piece, using Claude to enforce rules and banning AI-sounding words like 'cadence' and 'optimize'.

Local MCP Memory System with Consolidation for AI Conversations
A developer built an MCP server that provides persistent local memory for AI clients, using Qwen 2.5-7B to consolidate conversations into structured knowledge documents every 6 hours. The system runs entirely on your hardware with semantic dedup, adaptive scoring, and FAISS vector search.

Modulus: Cross-repository knowledge orchestration for AI coding agents
Modulus is a desktop app that runs multiple AI coding agents with shared project memory across repositories. It solves cross-repo context problems by letting agents understand dependencies between different codebases without manual explanation.