Pokemon Showdown AI Agents Built with Free LLM APIs and Tool-Calling

✍️ OpenClawRadar📅 Published: May 1, 2026🔗 Source
Pokemon Showdown AI Agents Built with Free LLM APIs and Tool-Calling
Ad

A developer built a system where LLMs like Llama 3, Qwen, and Gemma autonomously play Pokémon Showdown battles. The agents analyze the full battle state each turn—type matchups, HP, weather, field conditions, revealed opponent info—and decide whether to attack or switch using structured tool calls.

Key Details

  • Routes everything through LiteLLM and exclusively uses models with free API tiers (Groq, Cerebras, OpenRouter, Google AI Studio).
  • Zero inference cost to run locally.
  • Two modes: Human vs. AI (play against the bot) and AI vs. AI (pit two models against each other).
  • Supports 15+ free models out of the box.
  • Full observability via Langfuse to see exact tool calls and reasoning per turn.
Ad

Architecture Highlights

The agent uses tool-calling to structure decisions—rather than simple prompt-response—raw battlefield data is fed into the LLM, which then selects attack or switch actions via predefined tool schemas. This allows reasoning about complex board states like type advantages and dynamic field effects.

GitHub Repo

Code and setup instructions: github.com/MohamedMostafa259/pokemon-ai-agent

📖 Read the full source: r/LocalLLaMA

Ad

👀 See Also

AGENTS-COLLECTION: 129 Claude Code Agents Organized in One Repository
Tools

AGENTS-COLLECTION: 129 Claude Code Agents Organized in One Repository

A developer has compiled 129 Claude Code agents into a single repository in ~/.claude/agents/ format, ready for installation with a simple copy command. The collection includes the full agency-agents system with 68 personality-driven agents across multiple disciplines, plus additional agents for multi-agent team workflows.

OpenClawRadar
Terminal-Based 3D Renderer Built with Multi-Agent Claude Code System
Tools

Terminal-Based 3D Renderer Built with Multi-Agent Claude Code System

A developer created tortuise, a pure terminal-based 3D renderer that displays Gaussian splats using Unicode and ASCII symbols, built over 3 days using 70-80 AI agents coordinated through a Claude Code setup with subagents inside subagents.

OpenClawRadar
Claude Code Skill Delegates Coding to Mistral/DeepSeek: 57M Tokens Saved, 90-100% Cost Reduction
Tools

Claude Code Skill Delegates Coding to Mistral/DeepSeek: 57M Tokens Saved, 90-100% Cost Reduction

A Claude Code skill called vibe-skill delegates low-level coding to cheap models like Mistral or DeepSeek while keeping Claude's planning. After 254 runs over 10 days, it saved 57M tokens and achieved 90-100% cost savings with 98% success rate.

OpenClawRadar
Understudy: A Teachable Desktop Agent That Learns Tasks by Demonstration
Tools

Understudy: A Teachable Desktop Agent That Learns Tasks by Demonstration

Understudy is a local-first desktop agent runtime that can operate GUI apps, browsers, shell tools, files, and messaging in one session. You demonstrate a task once, it records screen video and semantic events, extracts intent rather than coordinates, and turns it into a reusable skill.

OpenClawRadar