Two new models appear on OpenRouter, possibly DeepSeek V4 variants

Two new models have appeared on OpenRouter that may be trial versions of DeepSeek V4. The models are named healer-alpha and hunter-alpha, with descriptions suggesting one is a Lite version and the other appears to be a full-featured model.
Model Specifications
The full version reportedly has 1TB of parameters and 1M of context, which matches leaked information about DeepSeek V4. The Lite version is described as a lighter variant of the same model family.
Initial Testing Results
A user conducted roleplay tests to evaluate filtering levels and performance:
- Both models performed impressively in roleplay scenarios
- Neither model declined any messages during testing
- The Lite version is noticeably faster than the full version
- The full version is slower but still responsive
- Both models generate the same amount of tokens in less than half the time compared to GLM 5.0
- The Lite version is slightly weaker in performance but not significantly
- Both models maintain character consistency and handle "spicy" content well
The models are currently in alpha phase, which may explain the lack of message filtering observed during testing. The community is discussing whether these are indeed DeepSeek V4 variants and sharing additional testing results.
📖 Read the full source: r/LocalLLaMA
👀 See Also

Claude MAX Plan Now Includes 1M Token Context Window at No Extra Cost
The Claude MAX plan has been automatically upgraded to include a 1 million token context window without additional API-based usage charges, with users reporting significantly reduced token usage and elimination of context window management overhead.

Autonoma's 18-month codebase rewrite: lessons on testing, tech debt, and Server Actions
Autonoma threw away 1.5 years of code after scaling from 2 to 14 engineers, citing no tests, unstrict TypeScript, and Server Actions limitations as key reasons for the rewrite.

OpenClaw v2026.3.11-beta.1 released with free AI models, cron breaking change
OpenClaw v2026.3.11-beta.1 introduces two free AI models on OpenRouter with 1M context windows, fixes Kimi coding tool calls, adds OpenCode provider support, and includes a breaking change for cron job notifications.

Nvidia RTX Spark: 1-Petaflop Superchip Brings Local AI Agents to Windows PCs
Nvidia unveils RTX Spark, a 1-petaflop superchip for Windows PCs, enabling local AI agents with up to 128GB unified memory and full CUDA/RTX stack. Ships this fall in laptops and desktops from ASUS, Dell, HP, Lenovo, Microsoft Surface, and MSI.