All skills
apify avatar

/apify-ultimate-scraper

@97047d3 official
by apifyapify/agent-skills2.4k stars
259

Universal AI-powered web scraper for any platform. Scrape data from Instagram, Facebook, TikTok, YouTube, LinkedIn, X/Twitter, Google Maps, Google Search, Google Trends, Reddit, Airbnb, Yelp, and 15+ more platforms. Use for lead generation, brand monitoring, competitor analysis, influencer discovery, trend research, content analytics, audience analysis, review analysis, SEO intelligence, recruitment, or any data extraction task.

Use this Skill: https://skilld.dev/gh/apify/agent-skills/apify-ultimate-scraper

This session only. Nothing lands on disk.

referencesworkflowscompetitive-intel.md

≈902 tokens on demand. Your agent reads this file only when SKILL.md points to it.

Competitive intelligence workflows

Competitor ad monitoring

When: User wants to see competitor advertising creatives, targeting, or ad spend signals.

Pipeline

  1. Scrape ad library -> apify/facebook-ads-scraper
    • Key input: searchQuery (competitor name), country, adType, maxItems

Output fields

Step 1: adTitle, adBody, adCreativeUrl, startDate, pageInfo.name, platform

Gotcha

Facebook Ad Library is public data, no auth needed. But results are limited to currently active or recently inactive ads.


Competitor web presence analysis

When: User wants traffic, rankings, and SEO data for competitor domains.

Pipeline

  1. Get traffic data -> radeance/similarweb-scraper
    • Key input: urls (competitor domains)
  2. Get backlink profile -> radeance/ahrefs-scraper
    • Key input: urls (same domains)

Output fields

Step 1: globalRank, monthlyVisits, bounceRate, avgVisitDuration, trafficSources Step 2: domainRating, backlinks, referringDomains, organicKeywords

Cost estimate

radeance/ Actors cost $0.005-0.0275/result. A single domain audit across both steps costs ~$0.04-0.06.


Competitor website change detection

When: User wants to monitor competitor pricing pages, feature announcements, or product pages and get alerted when meaningful changes occur.

Pipeline

  1. Detect changes -> tri_angle/website-changes-detector
    • Key input: startUrls (competitor page URLs), notificationEmail, checkIntervalHours

Output fields

Step 1: url, changedAt, diff (text diff), screenshotUrl

Gotcha

tri_angle/website-changes-detector handles baseline storage internally - do not attempt to manage baselines externally or you will lose the diff history between runs.


Competitor SERP position monitoring

When: User wants to track where competitor domains rank for target keywords and get alerted on significant position shifts.

Pipeline

  1. Scrape SERP rankings -> apify/google-search-scraper
    • Key input: queries (target keywords array), countryCode, maxResultsPerPage
  2. Track traffic estimates -> radeance/similarweb-scraper
    • Key input: urls (competitor domains)
    • Pipe: run separately per competitor domain after extracting domains from step 1 results

Output fields

Step 1: organicResults[].url, organicResults[].position, organicResults[].title Step 2: globalRank, monthlyVisits, trafficSources

Cost estimate

radeance/similarweb-scraper costs ~$0.02-0.03 per domain. For 5 competitors, budget ~$0.10-0.15 per weekly run.


Competitor feature and pricing benchmarking

When: User wants a structured comparison of competitor pricing tiers, feature lists, and positioning across 5-10 competitor sites.

Pipeline

  1. Crawl pricing and feature pages -> apify/website-content-crawler
    • Key input: startUrls (competitor pricing page URLs), maxCrawlDepth (set to 1), includeUrlGlobs
  2. Extract structured data -> AI node (GPT-4o or Claude)
    • Pipe: results[].text -> extraction prompt per competitor
    • Key input: extraction schema (tiers, prices, key features, positioning statement)

Output fields

Step 1: text (clean markdown with pricing tables), url, metadata.title Step 2: AI-extracted structured JSON with tiers, prices, feature flags per competitor

Gotcha

Set maxCrawlDepth: 1 and use includeUrlGlobs to restrict crawl to pricing and features paths only. Without this, WCC will crawl the full site and inflate cost significantly.

Source: SKILL.md on GitHub

1 alert16d5 checks · Risk SAFE
  • Gen Agent Trust Hub16d

    The skill provides comprehensive instructions for orchestrating and interacting with web scrapers (Actors) via the Apify CLI. It provides structured playbooks for different data extraction workflows (B2B lead generation, brand monitoring, social media analytics, etc.) and contains standard guidelines on configuration, parameter passing, and error handling. No malicious behaviors, obfuscation techniques, or hidden actions were detected.

  • Socket16d

    No alerts

  • Snyk16d

    Risk: MEDIUM · 1 issue

  • Runlayer7mo

    2/2 files flagged

  • ZeroLeaks5mo

    Score: 93/100 · 2 sections analyzed

Signed by skilld at 97047d3. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub yesterday.

Activeupdated 3 months ago
  • SEO
  • apify
  • scraping
  • web-scraping
  • instagram
  • facebook
  • tiktok
  • youtube
  • linkedin
  • twitter
  • google-maps
  • reddit
  • lead-generation
  • competitor-analysis
  • brand-monitoring

README badge

README badge for apify/agent-skills/apify-ultimate-scraper

Scrapes data from 100+ Apify Actors covering Instagram, Facebook, TikTok, YouTube, LinkedIn, X/Twitter, Google Maps, Reddit, Airbnb, Yelp, and 15+ other platforms via the Apify CLI. Use for lead generation, competitor analysis, brand monitoring, influencer discovery, review analysis, or SEO intelligence on any public web platform.

Generated from the current SKILL.md.

Does this skill work with all platforms?
It covers ~100 Actors across 15+ platforms including Instagram, Facebook, TikTok, YouTube, LinkedIn, X/Twitter, Google Maps, Reddit, Airbnb, Yelp, and more. Not all platforms are equally supported; check the actor-index.md reference for your target platform.
What authentication is required?
You need an Apify account and CLI token. Authenticate via `apify login` (OAuth), set `APIFY_TOKEN` as an environment variable, or source from a .env file.
Do I need to install anything locally?
Yes, Apify CLI v1.5.0 or later must be installed via npm (`npm install -g apify-cli`).
Can I export results in different formats?
Yes. Results can be fetched as JSON or CSV using `apify datasets get-items` with the `--format` flag.
What should I read before running a scraper for the first time?
Check `references/actor-index.md` to find the right Actor for your platform, then read `references/gotchas.md` for common pitfalls specific to that Actor before running.

Generated from the current SKILL.md. These answers refresh after source changes.