browser-automation

Vision-driven browser automation using Midscene. Operates from screenshots — no DOM or accessibility labels needed. Runs in headless Puppeteer — does NOT take over the user's mouse or keyboard. Also supports CDP mode and Bridge mode to connect to an existing Chrome. Use this skill when the user want

By web-infra-dev · 5,063 installs

npx skills add web-infra-dev/midscene-skills --skill browser-automation

Source repository · Upstream listing