AG

agent-browser-skill

Provides tools for browser automation, allowing AI agents to interact with web pages and capture UI state.

Install

mkdir -p .claude/skills/agent-browser-skill && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/3575" && unzip -o skill.zip -d .claude/skills/agent-browser-skill && rm skill.zip

Installs to .claude/skills/agent-browser-skill

Activation

This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.

基于 agent-browser CLI 的浏览器自动化工具。提供快照获取、元素交互、截图等功能。推荐用于需要页面快照分析、通过 ref 引用交互元素的场景。
79 charsno explicit “when” trigger
Intermediate

Key capabilities

  • →Get core workflow and command guides for agent-browser
  • →Get full command parameters and script templates for agent-browser
  • →Obtain dedicated skill guides for specific web pages
  • →Automate Electron desktop applications
  • →Automate Slack workspace operations
  • →Save screenshots using a system-injected environment variable

How it works

This skill provides access to the agent-browser CLI's documentation and specialized guides for browser automation tasks. It directs users to dynamically load current instructions for various scenarios.

Inputs & outputs

You give it
User request for browser automation, specific skill guides, or a screenshot command
You get back
Latest detailed workflow and command guides, dedicated skill guides, or a saved screenshot

When to use agent-browser-skill

  • →Capturing browser snapshots for debugging
  • →Automating interaction with a web application
  • →Extracting data from non-standard web interfaces

About agent-browser-skill

This skill uses the agent-browser CLI to interact with web applications, take snapshots, and perform automated tasks. It is ideal for exploratory testing, bug reporting, and interacting with web-based interfaces.

基于 agent-browser CLI 的浏览器自动化工具。提供快照获取、元素交互、截图等功能。推荐用于需要页面快照分析、通过 ref 引用交互元素的场景。

Prerequisites

agent-browser CLI

Limitations

  • →This file does not serve as the primary usage guide.
  • →Users must dynamically load the latest detailed workflow and command guides.
  • →Screenshot paths must use the SCREENSHOT_DIR environment variable.

How it compares

This skill provides dynamic, up-to-date guidance for a specific CLI tool, contrasting with static, potentially outdated documentation.

Compared to similar skills

agent-browser-skill side by side with the closest alternatives in the catalog.

SkillInstallsUpdatedSafetyDifficulty
agent-browser-skill (this skill)44moReviewIntermediate
dev-browser536moReviewIntermediate
agent-browser304moReviewIntermediate
browser-tools610moReviewIntermediate

Try saying

Example prompts that trigger this skill in your AI assistant.

You might also like

dev-browser

SawyerHood

Browser automation with persistent page state. Use when users ask to navigate websites, fill forms, take screenshots, extract web data, test web apps, or automate browser workflows. Trigger phrases include "go to [url]", "click on", "fill out the form", "take a screenshot", "scrape", "automate", "test the website", "log into", or any browser interaction request.

53176

agent-browser

vercel-labs

Browser automation CLI for AI agents. Use when the user needs to interact with websites, including navigating pages, filling forms, clicking buttons, taking screenshots, extracting data, testing web apps, or automating any browser task. Triggers include requests to "open a website", "fill out a form", "click a button", "take a screenshot", "scrape data from a page", "test this web app", "login to a site", "automate browser actions", or any task requiring programmatic web interaction.

3075

browser-tools

Whamp

Lightweight Chrome automation toolkit with shared configuration, JSON-first output, and six focused scripts for starting, navigating, inspecting, capturing, evaluating, and cleaning up browser sessions.

694

browser

cexll

This skill should be used for browser automation tasks using Chrome DevTools Protocol (CDP). Triggers when users need to launch Chrome with remote debugging, navigate pages, execute JavaScript in browser context, capture screenshots, or interactively select DOM elements. No MCP server required.

346

browserwing-executor

browserwing

Control browser automation through HTTP API. Supports page navigation, element interaction (click, type, select), data extraction, accessibility snapshot analysis, screenshot, JavaScript execution, and batch operations.

27

go-rod-master

rootcastleco

Comprehensive guide for browser automation and web scraping with go-rod (Chrome DevTools Protocol) including stealth anti-bot-detection patterns.

00

Search skills

Search the agent skills registry