Claude Pro User Reports 5-Hour Usage Window Burned on Single Prompt with No Output

A Reddit user /u/TaleOfACat reports a frustrating experience with Claude Pro: a single prompt consumed their entire 5-hour usage window (100%), yet the model returned no usable output. Instead of delivering code, Claude entered what the user calls “project architect mode” — outputting paragraphs like “surveyed project scope” and “architected comprehensive redesign” with no final file. The model then hit its own internal limit and prompted the user to send a “continue” message to proceed.
The user acknowledges that tokens are consumed as generation starts, but argues that as a paying user, this is a broken experience. Key complaints:
- Burning an entire usage window on internal “planning” that produces no deliverable.
- Encouraging another “continue” message after the usage is already drained.
- Lack of a goodwill adjustment or safety mechanism when the model clearly fails to deliver the requested output.
The post reflects a common pain point for Claude Pro users: the unpredictability of token consumption when the model engages in extensive internal reasoning without producing a final result. The user suggests that the product should at minimum avoid consuming the window on non-deliverable text and not prompt additional messages after draining the usage.
📖 Read the full source: r/ClaudeAI
👀 See Also

Top AI Models Show Performance Gap in Non-English Languages
A recent analysis shows leading AI models perform worse in languages other than English, with the article receiving 16 points and 3 comments on Hacker News.

Granite 4.1: IBM's 8B Dense Model Matches 32B MoE in Benchmarks
IBM's Granite 4.1 8B dense model matches or beats the previous 32B MoE model on ArenaHard, BFCL V3, GSM8K, and more, thanks to improved training data quality.

Apple Silicon Benchmark: Qwen3-VL Performance on M3, M4, and M5 Max for Vision LLM Classification
Benchmark results show Qwen3-VL vision LLM classification performance on Apple Silicon: M3 Max and M4 Studio are nearly identical for 8B models, while M5 Max is 75-83% faster. Memory bandwidth matters more for token generation than prefill in vision tasks.

Anthropic Removes Claude Code from Pro Subscription for New Users in Test
Anthropic temporarily removed access to Claude Code from its $20/month Pro subscription plan for new users, changing website pricing pages and support documents before reversing the changes. The company described it as a 'small test of 2% of new prosumer signups.'