Two new models appear on OpenRouter, possibly DeepSeek V4 variants

Two new models have appeared on OpenRouter that may be trial versions of DeepSeek V4. The models are named healer-alpha and hunter-alpha, with descriptions suggesting one is a Lite version and the other appears to be a full-featured model.
Model Specifications
The full version reportedly has 1TB of parameters and 1M of context, which matches leaked information about DeepSeek V4. The Lite version is described as a lighter variant of the same model family.
Initial Testing Results
A user conducted roleplay tests to evaluate filtering levels and performance:
- Both models performed impressively in roleplay scenarios
- Neither model declined any messages during testing
- The Lite version is noticeably faster than the full version
- The full version is slower but still responsive
- Both models generate the same amount of tokens in less than half the time compared to GLM 5.0
- The Lite version is slightly weaker in performance but not significantly
- Both models maintain character consistency and handle "spicy" content well
The models are currently in alpha phase, which may explain the lack of message filtering observed during testing. The community is discussing whether these are indeed DeepSeek V4 variants and sharing additional testing results.
📖 Read the full source: r/LocalLLaMA
👀 See Also

ICML 2026 Desk-Rejects 2% of Papers for LLM Review Policy Violations
ICML 2026 rejected 497 papers (~2% of submissions) after detecting 795 reviews (~1% of all reviews) where reviewers violated explicit agreements not to use LLMs. The detection method involved watermarking PDFs with hidden LLM instructions.

DeepSeek Withholds Latest AI Model from Nvidia and AMD
DeepSeek is withholding its latest AI model from U.S. chipmakers including Nvidia and AMD, according to Reuters sources. The article has 19 points and 3 comments on Hacker News.

AI Should Elevate Your Thinking, Not Replace It — Koshy John on the Hidden Divide in Engineering
Koshy John argues that engineers who outsource thinking to AI for short-term productivity gains are building a hollow foundation, while those who use AI to remove drudgery and operate at a higher level create real long-term value.

Claude Code v2.1.86: Session headers, memory fixes, and token optimizations
Claude Code v2.1.86 adds X-Claude-Code-Session-Id headers for proxy aggregation, fixes memory growth in long sessions, and reduces token overhead when mentioning files with @. The release addresses 18 specific issues including config corruption on Windows and OAuth URL copying.