Novice Builds Local 'Second Brain' on Intel Arc Pro B60: Qwen 3.8 27B at 38 tok/s, 128K Context
A non-developer with no coding background built a mostly local personal AI agent (OpenClaw) on an old Dell OptiPlex, then built a dedicated AI server around an Intel Arc Pro B60 GPU. The source is a detailed writeup of the exact hardware, costs, and results — including running Qwen 3.8 27B at ~38 tok/s with a 128K context window.
How it started
The author was an average chatbot user — research and writing only — until a YouTube video about OpenClaw (an AI agent that actually does things, not just chats) got them building. Setup happened in about 2 days using a free Claude account on an old Dell OptiPlex 3431, initially running on a $20/month ChatGPT subscription.
The Dell (agent's control plane)
- CPU: Intel Core i9-9900 (upgraded from stock i5)
- RAM: 32GB DDR4-2666 (was 64GB, half moved to the AI box)
- GPU: NVIDIA Quadro P620, 2GB (stock, useless for AI)
- OS drive: 512GB NVMe (added)
- Storage: 2TB SATA SSD (original)
- OS: Windows 11 Pro
Runs OpenClaw (agent harness), memory, scheduling, Telegram (the chat interface), backups, and stores local models deployable to the AI box at any time.
Stability issues traced back to cheap "grey market" memory sticks the Dell shipped with on Amazon — the author suspects bad RAM corrupted Windows files. The i9 upgrade was called out as the only bad spend so far ($200 on eBay), with the agent literally pleading against buying it.
The BlackBox (BB) — dedicated AI server
While running the agent, the author ordered then cancelled a $2,000 Mac Studio after realizing it couldn't run a model capable of doing what they needed. They briefly went all-cloud: $20/month Claude plus a $100 ChatGPT plan = $120/month just to run the agent, which felt like a trap.
The turning point was small open models — Qwen, GPT-OSS 20B, Mistral, Gemma. When Qwen 3.8 27B dropped, they committed to building a separate machine that does nothing but serve AI to the agent. That separation is described as the best decision of the project.
The guiding rule
Everything must be generic and interchangeable — never married to a model, a piece of hardware, or even the agent harness itself. That lets the agent host machine change while local inference stays put.
Reported numbers
Qwen 3.8 27B at roughly 38 tok/s with a 128K context window on the Intel Arc Pro B60 — now the default model, with cloud models as backup instead of the other way around. Full specs, config, and numbers are in the original post.
📖 Read the full source: r/openclaw
👀 See Also

Multi-Agent AI Teams Using Context Baptism to Improve Code Reviews
A developer running 18 generations of AI agent teams discovered that agents who read letters and retrospectives from previous generations write significantly better code reviews than those who only read the code, calling this practice 'Context Baptism.'

Speculative Decoding Benchmarks on RTX 3090 with Qwen Models for HVAC Business Use
A developer tested speculative decoding on an RTX 3090 using Qwen models for an HVAC business Discord bot, achieving up to 279.9 tokens/sec with a 236% speedup using Qwen3-8B with a Qwen3-1.7B draft model.

Non-coder builds live MLB dashboard using Claude AI and Claude Code on GitHub Codespaces
A user with no coding experience used Claude chat and Claude Code on GitHub Codespaces to build a live MLB dashboard with injury reports, game scores, and team stats, deploying it to Vercel.

Practical OpenClaw Setup Patterns from Real-World Deployments
A Reddit user shares insights from setting up OpenClaw for 10+ non-technical users, revealing that successful deployments typically involve 1-2 messaging apps, 5-10 simple workflows, local Mac operation, and voice cloning as a key adoption driver.