What problem does it solve? AI agents need a reliable way to interact with websites and desktop apps—navigating pages, filling forms, clicking buttons, and extracting data—without heavy dependencies like Playwright or Puppeteer. ## Core Features & Use Cases - Browser Automation: Navigate pages, fill forms, click elements, take screenshots, and scrape data using Chrome/Chromium via CDP with accessibility-tree snapshots and compact element refs. - Specialized Skills: Automate Electron desktop apps (VS Code, Slack, Discord, Figma), manage Slack workspaces, run exploratory QA and dogfooding sessions, and operate cloud browsers in Vercel Sandbox or AWS Bedrock AgentCore. - Use Case: Ask your agent to log into a web app, walk through the signup flow, capture screenshots of each step, and report any UI bugs found during exploratory testing. ## Quick Start Install the CLI with npm i -g agent-browser, then ask your agent to open a website and take a screenshot using agent-browser.