Automate your Chrome by screenshots
/chrome-bridge-automation
/chrome-bridge-automation is a Claude Code skill in the Code section. It is maintained by Hainrixz and published under the MIT license. Controls your own Chrome, with its logins and cookies, through Midscene Bridge, acting from screenshots rather than the DOM.
Author's description
Vision-driven browser automation using Midscene Bridge mode. Operates entirely from screenshots — no DOM or accessibility labels required. Can interact with all visible elements on screen regardless of technology stack. This mode connects to the user's desktop Chrome browser via the Midscene Chrome Extension, preserving cookies, sessions, and login state. Use this skill when the user wants to: - Browse, navigate, or open web pages in the user's own Chrome browser - Interact with pages that require login sessions, cookies, or existing browser state - Scrape, extract, or collect data from websites using the user's real browser - Fill out forms, click buttons, or interact with web elements - Verify, validate, or test frontend UI behavior - Take screenshots of web pages - Automate multi-step web workflows - Check website content or appearance Powered by Midscene.js (https://midscenejs.com)
How to ask for it
Open my Gmail in Chrome and find this week's unread emailsLog into my admin dashboard and fill out the new product formTake screenshots of my site after signing into my real account
Install
curl -fsSL https://raw.githubusercontent.com/sgomez-dev/claude-skills/main/install.sh | bashAfter installing with the script, type /chrome-bridge-automation. Using Cursor, Windsurf or Codex? Platform guides
Demo coming soon