apify

Run Apify cloud actors to scrape web content and retrieve structured datasets.

1|Updated Mar 11, 2026
One-click install
npx skills add https://github.com/antonyfmunoz/OS --skill apify-antonyfmunoz
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: apify
Source: https://github.com/antonyfmunoz/OS/tree/main/skills/tools/apify
Command: npx skills add https://github.com/antonyfmunoz/OS --skill apify-antonyfmunoz

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

Apify lets you reliably scrape websites and extract structured data when direct APIs or public endpoints are unavailable, especially for high-noise sources like Instagram.

Core Features & Use Cases

  • Run hosted scraping actors: Start pre-built or custom Apify “actors” to scrape targets (e.g., Instagram hashtags, comments, profiles) and collect results.
  • Retrieve structured datasets: Fetch dataset items from completed runs for downstream filtering and lead qualification.
  • Use proxy infrastructure: Optionally route traffic through Apify proxies (including residential sessions) to reduce blocking and access login-gated surfaces.
  • Apify-backed EOS pipeline: Powers Instagram lead-signal harvesting with bot/spam filtering, ICP relevance checks, and priority signal classification.

Quick Start

Use the apify skill to extract Instagram leads by running the configured hashtag and comment actors and saving qualified signals into EOS raw signal files.

Frequently Asked Questions about apify

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I scrape Instagram hashtags and profiles for lead generation?

Web scraping Instagram hashtags and profiles for lead generation uses Apify cloud actors to extract structured data like comments and user profiles. Retrieved datasets are then filtered for lead signal harvesting and competitor monitoring.

Can I use proxy sessions to access login-gated surfaces when web scraping?

Proxy sessions enable access to login-gated surfaces during web scraping by routing traffic through Apify proxies. Optional residential proxy credentials reduce blocking risks and support DM and login monitoring across batch cron-style runs.

Do I need an Apify API token to retrieve datasets from actor runs?

Retrieving datasets from actor runs requires token-based authentication using an Apify API v2 token. The skill executes actor runs and polls for completion to fetch structured dataset items for downstream lead qualification.

How does automated web scraping handle rate limiting and polling errors?

Automated web scraping handles rate limiting and polling errors through EOS-aligned rate limiting and built-in error handling during actor-run execution. This ensures reliable dataset retrieval across batch cron-style runs without manual intervention.

What is the best way to monitor competitor activity using web scraping?

Monitoring competitor activity using web scraping is best achieved by running Apify actors on a cron schedule to extract structured datasets. This approach harvests lead signals and tracks competitor changes reliably when direct APIs are unavailable.

What are the limitations of using cloud actors for web scraping high-noise sources?

Limitations of using cloud actors for web scraping high-noise sources like Instagram include potential blocking without proxy infrastructure and the need for bot and spam filtering. EOS pipelines classify priority signals to manage dataset noise.