HN data confirms arXiv paper share dropping, LLM hype peak behind us?

Dylan Castillo wanted to confirm whether he was seeing fewer arXiv papers on Hacker News front pages, so he used Claude to run a quick analysis against the BigQuery HN dataset. The results show a clear trend: the share of arXiv stories on HN has been declining sharply in the last few months.
He also looked at historical peaks. The first peak in 2019 was driven by deep learning papers — 41% of the top 100 upvoted arXiv posts that year were about deep learning. The 2023–2026 period saw an even heavier AI focus: 59% of the top 100 upvoted arXiv stories were about LLMs or AI. In 2019 the standout papers included MuZero (161 pts), EfficientNet (119 pts), XLNet (79 pts), the PyTorch NeurIPS paper (113 pts), and Chollet's “On the Measure of Intelligence” (80 pts).
For the 2023–2026 period, Castillo asked Claude to guess which papers will age well. The picks: DeepSeek-R1 (1,351 pts, open recipe for o1-style reasoning via RL), Generative Agents (391 pts, the “Smallville” paper), The Era of 1-bit LLMs / BitNet b1.58 (1,040 pts), Differential Transformer (562 pts), and the LK-99 cluster (2,408 + 1,690 pts combined, a landmark in open-science replication). The full analysis includes charts for topic distribution and the arXiv share over time.
📖 Read the full source: HN LLM Tools
👀 See Also

Kimi $19/m Update: Enhancing OpenClaw with Structured Models
Kimi introduces its latest update priced at $19/month, focusing on enhancing model structuring within OpenClaw. This update promises streamlined operations and improved automation features.

Mistral's Open-Weight Strategy: $14B Valuation on Sovereignty, Not Benchmarks
Mistral built a $14B AI empire by offering open-weight models for governments and enterprises seeking AI independence from US and Chinese tech. Revenue hit $200M in 2025, targeting $80M/month by Dec 2026.

Analysis of 100M tokens in Claude Code reveals 99.4% input usage
Analysis of 1,289 requests across extended coding sessions shows Claude Code used 100.3M input tokens (99.4%) versus only 616K output tokens (0.6%), with 84.2M tokens cached due to repeated context re-sending.

Mercury 2: Diffusion-Based Model for Real-Time AI Coding
Mercury 2 uses diffusion-based generation instead of sequential token-by-token decoding, generates tokens in parallel and refines them over steps, and claims 1,009 tokens/sec on NVIDIA Blackwell GPUs with pricing at $0.25/1M input tokens and $0.75/1M output tokens.