Claw-Written Plugin Adds Qwen 3.8 27B Thinking Level Support to OpenClaw
Running Qwen 3.8 27B on llama.cpp with OpenClaw 2026.6.9 but missing native thinking-level control? One OpenClaw user hit the same wall and let their AI agent write a plugin to fix it — fully generated, untested by human eyes.
Why You Need This Plugin
Qwen 3.8 controls thinking levels via the chat template rather than a dedicated parameter. OpenClaw's llama.cpp backend at version 2026.6.9 doesn't map this out of the box, so thinking levels get ignored or misapplied. The plugin bridges that gap.
What It Does
- Per-API-call thinking levels — correctly injects the
enable_thinkingvalue into the chat template for each request. - Separate level for heartbeats — set a different thinking level for background heartbeat calls (the author uses
xhigh). - Instruct mode (thinking off) — disables thinking and applies sampling parameters recommended by unsloth for optimal output.
- Optional timeout guard — force compaction to instruct mode to avoid timeouts during long reasoning chains.
How to Get It
The plugin is published on GitLab:
git clone https://gitlab.com/moltwithhat/llamacpp-qwen-thinking
Author's note: “I didn't even look at the code” — the entire thing was written by the AI agent itself. It's a proof-of-concept that agents can produce useful, shareable tools without human review.
Who It's For
Anyone running Qwen 3.8 (or similar models) on llama.cpp through OpenClaw who wants fine-grained control over thinking levels per request. If you rely on heartbeats or need instruct mode as a fallback, this saves you the plumbing work.
📖 Read the full source: r/openclaw
👀 See Also
Atlas: World Labs' Omni World Model for Spatial Intelligence
World Labs introduces Atlas, an omni world model that natively handles text, images, video, and 3D. It enables camera-controlled generation, spatial reconstruction, and space-time simulation.

Cortex v1.2 adds LLM enrichment, Q&A with citations, and conflict resolution
Cortex, a local memory layer for OpenClaw agents, has released v1.2 with LLM-augmented enrichment by default, a question-answering command with citations, and improved deduplication and conflict resolution. The tool now includes unified configuration management and intent-based search pre-filtering.

agentmemory V4 achieves 96.2% on LongMemEval benchmark, outperforms commercial AI memory systems
agentmemory V4 scored 96.2% on LongMemEval, beating several funded AI memory companies including PwC Chronos (95.6%), Mastra (94.87%), and OMEGA (93.2%). The system was built solo in 16 days on a mid-range gaming PC with a $1,000 budget.

Exporting AI Agent Memories Using Claude's Import Function
A Reddit user shares a prompt for extracting stored memories from AI agents like ChatGPT and Claude, then importing them into OpenClaw. The prompt requests all stored context including instructions, personal details, projects, tools, and preferences.