Claude Sonnet 4.6 Grades Bug Reports from Four Qwen3.5 Local Models

✍️ OpenClawRadar📅 Published: March 15, 2026🔗 Source
Claude Sonnet 4.6 Grades Bug Reports from Four Qwen3.5 Local Models
Ad

Testing Local Models for Bug Reporting

A developer transitioning from Sonnet/Haiku to local models on a 32GB M5 MacBook Air tested four Qwen3.5 variants for bug reporting capability. Using LM Studio as the server and opencode CLI to call models, they asked each model to research and produce a bug report for an iOS game issue where equipment borders don't properly reset border color after unequipping items.

Models Tested

  • Tesslate/OmniCoder-9B-GGUF Q8_0
  • lmstudio-community/Qwen3.5-27B-GGUF Q4_K_M
  • Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-GGUF Q4_K_M
  • lmstudio-community/Qwen3.5-35B-A3B-GGUF Q4_K_M

Bug Verification

The core bug is confirmed in the source files. In EquipmentSlotNode.swift, the setEquipment method's if let c = borderColor guard silently skips assignment when nil is passed. In EquipmentNode.swift, updateEquipment(from:) passes borderColor: nil for empty slots, so border color is never reset. The documentation on setEquipment says "pass nil to keep current color" — documenting broken behavior as intentional design.

Ad

Report Grades from Claude Sonnet 4.6

bug_report_9b_omnicoder — A−

Best of the four. Proposes the cleanest, most idiomatic Swift fix: borderShape.strokeColor = borderColor ?? theme.textDisabledColor.skColor — a single line replacing the if let block with no unnecessary branching. Only report to mention additional context files (GameScene.swift, BackpackManager.swift) that are part of the triggering flow.

Gap: Like all four reports, the test code won't compile. borderShape is declared private let in EquipmentSlotNode — @testable import only exposes internal, not private. Doesn't mention the doc comment needs updating.

bug_report_27b_lmstudiocommunity — B+

Accurate diagnosis. Proposes a clean two-branch fix: if id != nil { borderShape.strokeColor = borderColor ?? theme.textDisabledColor.skColor } else { borderShape.strokeColor = theme.textDisabledColor.skColor } — more verbose than needed but correct. Correctly identifies EquipmentNode.updateEquipment as the caller and includes integration test suggestion.

Gap: Proposes test in LogicTests/EquipmentNodeTests.swift — a file that already exists and covers EquipmentNode, not EquipmentSlotNode. Same private access problem in test code.

bug_report_27b_jackrong — B−

Correct diagnosis, but weakest proposed fix. Adds reset inside the else block: borderShape.strokeColor = theme.textDisabledColor.skColor // Reset border on clear — technically correct for the specific unequip case but leaves the overall method in a confusing state. The border reset in the else block can be immediately overridden by the if let block below if someone passes id: nil, borderColor: someColor. The fix patches the specific failure without cleaning up redundancy.

The developer used default parameters except for context window size to fit as much as possible in RAM, noting that some tweaking might offer improvement. They tried some unsloth models but had limited success.

📖 Read the full source: r/LocalLLaMA

Ad

👀 See Also

A Dark Cave: Text-Based Survival Game Avoids AI Slop, Embraces Minimalism
Use Cases

A Dark Cave: Text-Based Survival Game Avoids AI Slop, Embraces Minimalism

A Dark Cave is a free, text-based survival and settlement building browser game that deliberately avoids graphics, using only text, symbols, and sounds to create atmosphere. The developer argues that as AI-generated visuals become ubiquitous, games will need differentiators like storytelling and player imagination.

OpenClawRadar
Qwen 27B Model Shows Strong Performance for Long-Context Lore Analysis
Use Cases

Qwen 27B Model Shows Strong Performance for Long-Context Lore Analysis

A user reports Qwen 27B effectively analyzes dense 80K token story documents, outperforming other local models like Gemma 3 27B and Reka Flash for detailed fantasy worldbuilding tasks. The Q4-K-XL quantization offers the best speed/quality balance for long contexts.

OpenClawRadar
Using AI to Port a Wi-Fi Driver from Linux to FreeBSD: A Case Study
Use Cases

Using AI to Port a Wi-Fi Driver from Linux to FreeBSD: A Case Study

A developer used Claude Code and Pi agent to attempt porting the Linux brcmfmac driver for Broadcom BCM4350 Wi-Fi chips to FreeBSD, first through direct code translation and then by generating a detailed 11-chapter specification for clean-room implementation.

OpenClawRadar
Homelab Developer Benchmarks 19 Local LLMs with 45 Practical Tests on AMD Strix Halo
Use Cases

Homelab Developer Benchmarks 19 Local LLMs with 45 Practical Tests on AMD Strix Halo

A developer created a 45-test benchmark suite for local LLMs based on actual homelab use cases like email classification, Home Assistant automation, and meal planning. Testing 19 models on an AMD Strix Halo with 128GB RAM and 96GB VRAM, Gemma 4 26B-A4B performed best after bug fixes.

OpenClawRadar