ProofShot: CLI for AI Agents to Verify UI Code with Browser Recording

What ProofShot Does
ProofShot is a CLI tool that gives AI coding agents visual verification capabilities. It allows agents to see what the UI they build actually looks like in the browser, detect layout issues, and capture console errors.
How It Works
The tool operates through three main commands:
proofshot start --run "npm run dev" --port 3000- Launches your dev server, opens headless Chromium, and starts recording video- Your AI agent then executes actions like
proofshot exec navigate "http://localhost:3000"andproofshot exec screenshot "homepage"to navigate, click, fill forms, and take screenshots proofshot stop- Collects errors, stops recording, trims dead time, and generates proof artifacts
Output and Features
ProofShot generates a standalone HTML file containing:
- Video playback of the browser session synced with an action timeline
- Screenshots taken during the session
- Element labels for each action
- Browser console errors captured during the session
- Server logs scanned with pattern matching for JavaScript, Python, Go, Rust, and other languages
- PR-ready artifacts including SUMMARY.md and formatted output for pull requests
- Visual diff comparison against baselines
Technical Details
The tool is:
- Built on agent-browser from Vercel Labs (described as "far better and faster than Playwright MCP")
- Not a testing framework - the agent doesn't decide pass/fail, it just provides evidence
- Agent-agnostic - works with Claude Code, Cursor, Codex, Gemini CLI, Windsurf, and any MCP-compatible agent
- Packaged as a skill so AI agents know exactly how it works
- Open source with MIT license
Installation and Setup
$ npm install -g proofshot
$ proofshot install
The tool automatically trims dead time from recordings, so you see only what the agent actually did, not idle waiting periods.
📖 Read the full source: HN LLM Tools
👀 See Also

Culpa: Open Source Deterministic Replay Engine for AI Agent Debugging
Culpa is an open source tool that records LLM agent sessions with full execution context, enabling deterministic replay using recorded responses as stubs instead of hitting real APIs. It works with Anthropic and OpenAI APIs via proxy mode or Python SDK.

SlackClaw: Managed OpenClaw Instance for Slack Integration
SlackClaw is a commercial product built on OpenClaw that provides a managed instance specifically for Slack. It offers one-click installation, OAuth tool connections, dedicated servers per workspace, and persistent memory.
Zillow-Full: An OpenClaw Skill That Turned Manual Property Research Into an Automated Deal Pipeline
A developer built 'zillow-full' on OpenClaw to pull Zestimates, tax history, price history, and comps per property. With a nightly cron scoring listings against deal criteria, wholesale deals went from 2 to 11 per month.

OpenClaw A2A Plugin: Direct Agent-to-Agent Messaging Over the Internet
An OpenClaw A2A plugin enables direct file and message transfer between OpenClaws and other agents over the internet without third-party services like WhatsApp or email.