Claude Code v2.1.201 Drops Mid-Conversation System Role for Sonnet 5 Sessions

Anthropic pushed Claude Code v2.1.201 on July 3rd, with a single but meaningful change: Claude Sonnet 5 sessions no longer inject harness reminders via the mid-conversation system role.
What changed
Previously, during a Sonnet 5 session, the harness (the tool-calling infrastructure) would periodically re-insert system-level reminders about available tools, permissions, and context rules into the conversation. This consumed tokens and could clutter the chat history. The v2.1.201 release removes that behavior entirely for Sonnet 5 sessions.
The change is specific to Claude Sonnet 5 — other models are unaffected. The exact commit is c489eb2.
Why this matters
For developers running long coding sessions with Claude Code, the mid-conversation system role meant every turn had hidden overhead in both token count and prompt clarity. Removing it should reduce token spend and make the conversation thread cleaner, especially for iterative tasks like debugging or code generation where the harness context doesn't need re-explaining.
If you're using Claude Code with Sonnet 5, update to this release to avoid unnecessary system-role inserts. Agents that relied on these reminders to stay on track (unlikely, since the reminders are for the harness, not the model's reasoning) should verify behavior after the update.
📖 Read the full source: GitHub Claude-Code
👀 See Also

DeepSeek-V4 Pro and Flash: 1.6T Parameters, 1M Token Context, Hybrid Attention
DeepSeek-V4-Pro (1.6T params, 49B active) and V4-Flash (284B params, 13B active) support 1M token context. New hybrid attention (CSA + HCA) reduces single-token inference FLOPs to 27% and KV cache to 10% of DeepSeek-V3.2.

Grammar-Based Method Matches or Outperforms AI in Authorship Analysis
A University of Manchester study found that LambdaG, a grammar-based authorship analysis method, matched or exceeded leading AI systems across most test datasets while offering greater transparency and lower computational cost.

Wikipedia Bans AI-Generated Content, Allows Limited AI Use with Human Review
Wikipedia has officially banned its 260,000 editors from using AI like ChatGPT to write articles, citing accuracy and reliability concerns. Editors can still use AI for translation and copy editing with human approval.

GitHub Copilot Moves to Usage-Based Pricing: The End of Subsidized AI Coding
Microsoft will charge GitHub Copilot users by actual model costs starting June 1, 2026, ending the $20+/month subsidy per user. Agentic AI usage is cited as the reason.