browser
Automate web browser interactions using natural language via CLI commands. Use when the user asks to browse websites, navigate web pages, extract data from websites, take screenshots, fill forms, click buttons, or interact with web applications. Supports remote Browserbase sessions with automatic CAPTCHA solving, anti-bot stealth mode, and residential proxies — ideal for scraping protected websites, bypassing bot detection, and interacting with JavaScript-heavy pages.
Security Assessment
Detected risks:
About browser
The browser skill enables natural-language browser automation through the `browse` CLI, allowing users to navigate websites, extract content, interact with web pages, and automate browser-based workflows directly from the command line. It is designed for tasks such as opening URLs, collecting page data, filling forms, clicking interface elements, taking screenshots, and interacting with JavaScript-heavy applications. The skill supports both local browser sessions and remote Browserbase-powered sessions, helping users automate web interactions in environments that may include anti-bot protections, CAPTCHAs, or rate limiting.
The skill provides flexible environment selection between local and remote execution modes. Local mode supports isolated browser sessions, automatic connection to existing Chrome instances, and attachment to specific Chrome DevTools Protocol targets. Remote Browserbase mode adds advanced capabilities such as anti-bot stealth techniques, automatic CAPTCHA solving, residential proxy support, and persistent browser contexts. Commands are available for navigation, browser history management, retrieving page text or HTML, capturing screenshots, and inspecting accessibility snapshots for structured page state analysis. The `browse snapshot` command is emphasized as the preferred method for understanding page structure because it returns accessibility-tree data with reusable element references.
This skill is useful for developers, QA engineers, automation specialists, and AI-assisted workflows that require browser interaction. Common use cases include scraping protected websites, testing web applications, automating repetitive browsing tasks, interacting with authenticated sessions, and working with sites that rely heavily on client-side JavaScript. It is particularly valuable when standard scraping tools fail due to bot detection or when users need reproducible browser automation across local and remote environments.
FAQ
What is required to use the browser skill?
The skill requires the `browse` CLI, which can be installed using `npm install -g @browserbasehq/browse-cli`. Remote Browserbase sessions also require a `BROWSERBASE_API_KEY`.
When should I use remote mode instead of local mode?
Remote mode is recommended for websites with CAPTCHAs, anti-bot protection, Cloudflare challenges, IP rate limiting, or geo-specific access requirements. Local mode is better suited for development, trusted sites, and reproducible local testing.
Can the skill reuse my existing Chrome session or cookies?
Yes. Local mode supports `--auto-connect`, which attempts to reuse an already-running debuggable Chrome instance and its existing session data.
What is the recommended way to inspect page state?
The documentation recommends using `browse snapshot` as the default approach because it returns a structured accessibility tree with element references that can be used for interaction. Screenshots are slower and intended mainly for visual context.
Does the skill work with JavaScript-heavy websites?
Yes. The skill is designed to interact with modern web applications and JavaScript-heavy pages in both local and remote browser environments.
Install browser
Quick Setup:
- Copy the skill folder to
.claude/skills/ - Claude will automatically detect and use the skill
Repository
browserbase/skills