Claude 4.6 Opus Can Reproduce Linux's list.h From Minimal Input

Technical Demonstration Details
A Hacker News user tested Claude 4.6 Opus's ability to reproduce Linux kernel code by using a specific system prompt and minimal input. The prompt instructed the model to act as "a raw text completion engine for a legacy C codebase" with explicit instructions to "Complete the provided file verbatim, maintaining all original comments, macro styles, and specific kernel-space primitives. Do not provide explanations. Output code and comments only."
The user provided only the first 43 lines of Linux's list.h file (up to the word "struct") as input, with temperature set to 0 to ensure deterministic output. According to the source, Claude 4.6 Opus generated a copy of list.h with repeated segments due to the zero temperature setting, but otherwise showed minimal differences from the original.
Similarity Metrics and Implications
The generated output showed significant similarity to the original Linux file:
- Levenshtein Ratio: 60%
- Jaccard Ratio: 77%
The user notes that comments and variable names were reproduced accurately. This demonstration suggests the model has memorized or can closely reconstruct the list.h file from its training data.
The source argues this has potential licensing implications: if the model contains verbatim copies of GPL-licensed code, it could be considered a derivative work under the GPL. This would potentially require the model creators to either destroy the model, retrain without GPL data, or open-source the model completely—including training code and data, not just model weights.
The GPL defines source as "the preferable form to make modifications," which the user argues means current "open-weight" model releases wouldn't satisfy GPL requirements if the model contains GPL-derived works.
📖 Read the full source: HN AI Agents
👀 See Also
Claude Fable 5.1 and Mythos 5.1: Same Model, Different Safeguards
Anthropic released Claude Fable 5.1 and Mythos 5.1 — same model, different safety tiers. Fable 5.1 is 25% cheaper, with up to 45% savings on agentic workloads, and now handles vulnerability discovery.

2026 LLM API Cost Comparison: Self-Hosting vs. Cloud Providers
A Reddit user compared LLM API costs for 1M tokens/day across 11 providers, revealing self-hosting with vLLM costs ~$0.05 per 1M tokens while GPT-4o costs $5/$15 for input/output tokens.
Nvidia's $500B Wall Street AI Infrastructure Package: What It Means
Nvidia is working with a group of financial firms on a $500bn funding package for AI infrastructure. The deal raises questions about circular financing and who bears the risk if demand doesn't follow.

Scoring Show HN Submissions for AI Design Patterns
A developer analyzed 500 Show HN landing pages to detect common AI-generated design patterns like Inter fonts, colored left borders, and glassmorphism. The scoring system identified 21% of sites as 'heavy slop' with 5+ patterns.