6.1k browser-act

browser-act Skill

面向 AI 智能体的浏览器自动化 CLI。切勿直接用 Bash 运行 browser-act 命令——务必先调用本技能。当用户提到 browser-act 或要求运行其 CLI 命令时使用,可用于:抓取、查看或提取 URL 渲染后的内容,访问需要 JavaScript 的页面,处理验证提示,维持登录会话,填表、点击走流程,输入、选择、上传、截图,捕获 XHR/fetch/HAR 响应,并行打开多个 URL,提取滚动或点击后才加载的内容。

安装方式:把技能目录放入 ~/.claude/skills/(Claude Code)或在 claude.ai 设置中启用;也可复制右侧安装命令一键添加。

查看源码

技能指令原文(SKILL.md)

browser-act

Built by BrowserAct — Browser automation CLI for AI agents · GitHub

Runs a full browser engine: navigation & interaction, data extraction & network
capture, screenshots, form automation, multi-browser parallel operation,
user-configured proxy support, and human-agent collaboration.

Features

  • Lightweight extraction — fast JS-rendered content fetch without opening a browser session, advanced WebFetch/curl replacement
  • Session management — multi-browser isolation, multi-account parallel operation
  • Verification assistance — when automation encounters interactive challenges, assists completion with user authorization
  • Complex interaction — DOM content extraction, screenshots, form filling, file upload
  • Human-agent collaboration — headed mode + remote assist for manual steps
  • Safety controls — Confirmation Gate protocol requires explicit user approval before browser creation, deletion, and sensitive operations
  • Universal compatibility — works with Cursor, Claude Code, Codex, Windsurf, etc.

Install: uv tool install browser-act-cli --python 3.12

Start here

This file is a discovery stub, not the usage guide. After loading this
skill, immediately run the following to get the actual workflow content:

browser-act get-skills core --skill-version 2.0.2

The CLI serves skill content that always matches the installed version,
so instructions never go stale. Do NOT truncate the output — none of
which are available through --help.