Testreel: Programmatic Demo Video Generation with Claude Code

What Testreel Does
Testreel is an npm package that lets you describe user interactions (click, type, scroll, zoom, etc.) in JSON, YAML, or Playwright format, then generates polished demo videos with visual effects.
Key Features from Source
- Generate webm, mp4, or gif videos from interaction descriptions
- Add cursor overlays, click ripples, and gradient backgrounds
- Use JSON, YAML, or Playwright to define interactions
- Integrated with Playwright for creating videos using mocks and sample data
- MIT License
- Built on Playwright + FFMPEG
Primary Use Cases
The source identifies two main value propositions:
- No need to manually re-record demos due to typos or misclicks - just update the config and regenerate
- Allows LLM agents (like Claude Code) to generate demo videos for web apps with cursor overlays and customizable desktop backgrounds
Practical Implementation
According to the source, users can ask Claude Code to use Testreel to generate demo videos of specific UI flows by describing what they want (e.g., "use realistic data," "use this image"). The author found this relatively straightforward to implement.
The tool is positioned as a programmatic version of ScreenStudio or Cap, enabling automated demo creation similar to how Playwright handles end-to-end testing.
Repository Information
The package is available at: https://github.com/greentfrapp/testreel
The author notes they created it with Claude Code and will continue updating it based on their own usage.
📖 Read the full source: r/ClaudeAI
👀 See Also

Open-source Claude Code skill diagnoses AI adoption roadblocks
An MIT-licensed Claude Code skill analyzes where companies get stuck with AI adoption—tooling, culture, or measurement—and builds 90-day plans with named owners. Based on interviews with 100+ founders and board members.

Detecting Silent Tool Failures in AI Coding Agents with Vibeyard
Vibeyard is a tool that detects when AI coding agents experience silent tool failures—where agents fall back to alternative strategies without alerting developers—and surfaces these inefficiencies during sessions. It can suggest fixes to prevent repeated inefficient workflows.

OpenClaw Model Performance Review: Codex 5.3 Leads, GLM Models Disappoint
A developer tested multiple AI models with OpenClaw, finding Codex 5.3 performs best with 9/10 rating, while GLM 4.7 and GLM 5 scored 5/10 due to high token usage, slow responses, and inconsistent output.

DIY OpenClaw Alternative Using Claude Code in Headless Mode
A developer built a Python server that sends prompts to Claude Code in headless mode, with Telegram bot access, Hammerspoon automation, and local markdown file storage for tasks, schedules, and notes.