Running Local LLM Agents on Mac Minis with Telegram Interface

A developer on r/LocalLLaMA detailed a system for running multiple local LLM agents on Mac Minis, controlled entirely through Telegram messages from a phone. The setup eliminates API costs and provides complete privacy while maintaining functionality similar to commercial services like Claude Code Channels.
Technical Setup
The core system uses:
- Local models through LMStudio: 35B models for everyday tasks, 235B models for heavier reasoning
- Claude Code running in tmux sessions on each Mac Mini
- Telegram bots that bridge user messages to the tmux sessions
- 80 lines of Python for the Telegram bot implementation (available on GitHub)
The workflow is straightforward: text a message to the Telegram bot, which types it into the tmux session, watches for output, and sends the response back.
Key Advantages
- Zero ongoing cost: Hardware is the only expense—no API keys, rate limits, or quota restrictions
- Complete privacy: Everything stays on the local area network (LAN)
- Model flexibility: Mix and match different models—one agent runs Gemini CLI, others use LMStudio pointed at Ollama models
- No vendor lock-in: LMStudio serves the Anthropic Messages API natively, so Claude Code connects to it as if talking to Anthropic's servers
Current Implementation
The developer runs 5 specialized agents, each with its own Telegram bot:
- Approval workflows with inline Telegram buttons (Approve/Reject/Tweak) for reviewing drafts from a phone
- Shared memory across agents via git synchronization
- Media generation (FLUX.1, Wan 2.2) dispatched to a GPU box
- Podcast pipeline with cloned voice TTS, triggered from a single Telegram message
Hardware Requirements
- 35B models: Run well on 64GB+ RAM Mac or 24GB GPU
- 235B models: Need 128-256GB RAM or multiple GPUs
- The developer recommends starting small and scaling as needed
The tmux bridge pattern is model-agnostic—it doesn't care what's running inside the session, allowing for easy swapping of underlying models. A full build guide for a single machine/agent is available, with multi-machine instructions coming soon.
📖 Read the full source: r/LocalLLaMA
👀 See Also
Building an OpenClaw Multi-Agent Assistant on Raspberry Pi 5
A dev shares the architecture for JDM Assistant: OpenClaw as the OS for a multi-agent personal assistant on Raspberry Pi 5, with failover and security gates.

Practical Limits of Multi-GPU AI Workstations: Lessons from a 9× RTX 3090 Build
A developer shares experience running 9 RTX 3090 GPUs for AI work, finding diminishing returns beyond 6 GPUs and recommending Proxmox for LLM experimentation. The RTX 3090 remains compelling at $750 for 24GB VRAM.

Building Drivesidekick: A Driving App with Claude Code
Developers are using Claude Code to build mobile apps without front-end expertise. A backend developer utilized Claude Code to create Drivesidekick, a driving lessons app utilizing React Native/Expo.

OpenClaw Cost Optimization: How a Developer Fixed a $750 Mistake with Model Routing
A developer shares how switching all OpenClaw subagents to the free Hunter Alpha model on OpenRouter led to silent failures, including a video production agent that generated valid code but produced a 9-second silent black video. The solution involved implementing explicit model routing based on task requirements.