What problem does it solve?
Manually browsing websites to complete repetitive tasks like form filling, data scraping, screenshot capture, or login flows is time-consuming, error-prone, and impossible to scale for large volumes of work.
Core Features & Use Cases
- Ref-Based Element Selection: Uses accessibility snapshot refs (@e1, @e2) for stable, reliable element targeting that avoids broken selectors from dynamic page content changes.
- Comprehensive Browser Actions: Supports navigation, form filling, clicking, scrolling, text/HTML extraction, full page or viewport screenshots, and PDF generation for any web page.
- Use Case: For example, you can use this skill to automatically log into a research database, search for papers matching your keywords, and export all result titles and URLs to a CSV file without manual clicking or copying.
Quick Start
Use the agent-browser skill to navigate to your target job board, fill the search bar with "remote software engineer", and scrape all matching job titles, companies, and application links into a structured list.