What problem does it solve?
This Skill eliminates the manual effort of interacting with websites by reliably performing browser-based actions and returning evidence like screenshots and extracted text.
Core Features & Use Cases
- Browser automation with interactive refs: navigate pages, snapshot interactive elements, then click/fill/select using stable element references.
- Evidence-based outputs: capture screenshots, PDF, page text, URLs, titles, and session states after real page execution.
- Authentication and session reuse: import auth from an existing Chrome session, persist state across runs, and reuse named sessions.
- Works across common UI patterns: form submission, multi-step flows, iframes, and lightweight page testing where DevTools-grade diagnostics are not required.
Quick Start
Ask your agent: Open the provided website, log in if needed, then click the requested button, and finally return a screenshot plus the extracted result text.