Jan-Code-4B: A Lightweight Code-Tuned Model for Local Development

Jan-Code-4B Release Details
The Jan team has released Jan-Code-4B, a small code-tuned model built on Jan-v3-4B-base-instruct. This experimental model targets day-to-day coding assistance tasks including code generation, edits/refactors, basic debugging, and writing tests while maintaining a lightweight footprint suitable for local execution.
Intended Use and Performance
Jan-Code-4B is designed as a drop-in replacement for the Haiku model in Claude Code. On coding benchmarks, it shows small improvements over the baseline model and generally feels more reliable for coding-oriented prompts at this size.
How to Run Jan-Code-4B
Setup via Jan Desktop:
- Download Jan Desktop from https://www.jan.ai/
- Download Jan-Code via Jan Hub
Claude Code Integration:
- Jan makes it easier to connect Claude Code to any model
- Replace Haiku model with Jan-Code-4B
Model Links and Parameters
Model downloads:
- Jan-Code: https://huggingface.co/janhq/Jan-code-4b
- Jan-Code-GGUF: https://huggingface.co/janhq/Jan-code-4b-gguf
Recommended parameters:
- Temperature: 0.7
- Top_p: 0.8
- Top_k: 20
The Jan team credits u/Alibaba_Qwen for the base model and u/ggerganov for llama.cpp contributions.
📖 Read the full source: r/LocalLLaMA
👀 See Also

MetaBot: Open-Source Bridge Connects Claude Code to Telegram, Feishu, and WeChat
MetaBot is an open-source TypeScript bridge that connects the Claude Code Agent SDK to messaging platforms like Telegram, Feishu, and WeChat. It provides persistent memory, scheduled tasks, multi-agent collaboration, and real-time streaming of tool calls.

Parallel Agent Orchestrator for Claude Code Using Git Worktrees
A developer built a parallel orchestrator that uses git worktrees to create isolated environments for Claude Code agents, solving the problem of shared working directories causing broken apps and messy git status.

yburn: Tool to audit and replace unnecessary AI agent cron jobs
yburn is a Python tool that audits AI agent cron jobs and replaces those that don't need LLMs with standalone Python scripts. The creator found 58% of 98 cron jobs were purely mechanical tasks like system health checks and git backups.

SubQ: A Sub-Quadratic LLM with 12M-Token Context Window
SubQ is a fully sub-quadratic sparse-attention LLM offering a 12M-token context window at 150 tokens/s, with SWE-Bench Verified 81.8% and RULER @ 128K 95.0%. It reduces attention compute ~1000× compared to transformers.