agent-browser-skill
Provides tools for browser automation, allowing AI agents to interact with web pages and capture UI state.
Install
mkdir -p .claude/skills/agent-browser-skill && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/3575" && unzip -o skill.zip -d .claude/skills/agent-browser-skill && rm skill.zipInstalls to .claude/skills/agent-browser-skill
Activation
This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.
基于 agent-browser CLI 的浏览器自动化工具。提供快照获取、元素交互、截图等功能。推荐用于需要页面快照分析、通过 ref 引用交互元素的场景。Key capabilities
- →Get core workflow and command guides for agent-browser
- →Get full command parameters and script templates for agent-browser
- →Obtain dedicated skill guides for specific web pages
- →Automate Electron desktop applications
- →Automate Slack workspace operations
- →Save screenshots using a system-injected environment variable
How it works
This skill provides access to the agent-browser CLI's documentation and specialized guides for browser automation tasks. It directs users to dynamically load current instructions for various scenarios.
Inputs & outputs
When to use agent-browser-skill
- →Capturing browser snapshots for debugging
- →Automating interaction with a web application
- →Extracting data from non-standard web interfaces
About agent-browser-skill
This skill uses the agent-browser CLI to interact with web applications, take snapshots, and perform automated tasks. It is ideal for exploratory testing, bug reporting, and interacting with web-based interfaces.
基于 agent-browser CLI 的浏览器自动化工具。提供快照获取、元素交互、截图等功能。推荐用于需要页面快照分析、通过 ref 引用交互元素的场景。
Prerequisites
Limitations
- →This file does not serve as the primary usage guide.
- →Users must dynamically load the latest detailed workflow and command guides.
- →Screenshot paths must use the SCREENSHOT_DIR environment variable.
How it compares
This skill provides dynamic, up-to-date guidance for a specific CLI tool, contrasting with static, potentially outdated documentation.
Compared to similar skills
agent-browser-skill side by side with the closest alternatives in the catalog.
| Skill | Installs | Updated | Safety | Difficulty |
|---|---|---|---|---|
| agent-browser-skill (this skill) | 4 | 4mo | Review | Intermediate |
| dev-browser | 53 | 6mo | Review | Intermediate |
| agent-browser | 30 | 4mo | Review | Intermediate |
| browser-tools | 6 | 10mo | Review | Intermediate |
Try saying
Example prompts that trigger this skill in your AI assistant.
More by MGdaasLab
View all by MGdaasLab →You might also like
dev-browser
SawyerHood
Browser automation with persistent page state. Use when users ask to navigate websites, fill forms, take screenshots, extract web data, test web apps, or automate browser workflows. Trigger phrases include "go to [url]", "click on", "fill out the form", "take a screenshot", "scrape", "automate", "test the website", "log into", or any browser interaction request.
agent-browser
vercel-labs
Browser automation CLI for AI agents. Use when the user needs to interact with websites, including navigating pages, filling forms, clicking buttons, taking screenshots, extracting data, testing web apps, or automating any browser task. Triggers include requests to "open a website", "fill out a form", "click a button", "take a screenshot", "scrape data from a page", "test this web app", "login to a site", "automate browser actions", or any task requiring programmatic web interaction.
browser-tools
Whamp
Lightweight Chrome automation toolkit with shared configuration, JSON-first output, and six focused scripts for starting, navigating, inspecting, capturing, evaluating, and cleaning up browser sessions.
browser
cexll
This skill should be used for browser automation tasks using Chrome DevTools Protocol (CDP). Triggers when users need to launch Chrome with remote debugging, navigate pages, execute JavaScript in browser context, capture screenshots, or interactively select DOM elements. No MCP server required.
browserwing-executor
browserwing
Control browser automation through HTTP API. Supports page navigation, element interaction (click, type, select), data extraction, accessibility snapshot analysis, screenshot, JavaScript execution, and batch operations.
go-rod-master
rootcastleco
Comprehensive guide for browser automation and web scraping with go-rod (Chrome DevTools Protocol) including stealth anti-bot-detection patterns.