All skills
apify avatar

/apify-ultimate-scraper

@97047d3 official
by apifyapify/agent-skills2.4k stars
259

Universal AI-powered web scraper for any platform. Scrape data from Instagram, Facebook, TikTok, YouTube, LinkedIn, X/Twitter, Google Maps, Google Search, Google Trends, Reddit, Airbnb, Yelp, and 15+ more platforms. Use for lead generation, brand monitoring, competitor analysis, influencer discovery, trend research, content analytics, audience analysis, review analysis, SEO intelligence, recruitment, or any data extraction task.

Use this Skill: https://skilld.dev/gh/apify/agent-skills/apify-ultimate-scraper

This session only. Nothing lands on disk.

referencesworkflowsreal-estate-and-hospitality.md

≈948 tokens on demand. Your agent reads this file only when SKILL.md points to it.

Real estate and hospitality workflows

Property search and analysis

When: User wants to find and compare property listings in a specific area.

Pipeline

  1. Search properties -> tri_angle/redfin-search
    • Key input: location, propertyType, minPrice, maxPrice
  2. Get details -> tri_angle/redfin-detail
    • Pipe: results[].url -> startUrls
    • Key input: startUrls

Output fields

Step 1: address, price, beds, baths, sqft, url, status Step 2: description, yearBuilt, lotSize, priceHistory[], taxHistory[], schools[]

Airbnb market analysis

When: User wants to analyze Airbnb listings, pricing, and reviews in a destination.

Pipeline

  1. Search listings -> tri_angle/new-fast-airbnb-scraper
    • Key input: location, checkIn, checkOut, maxItems
  2. Get reviews -> tri_angle/airbnb-reviews-scraper
    • Pipe: results[].url -> startUrls
    • Key input: startUrls, maxReviews

Output fields

Step 1: name, price, rating, reviews, type, amenities[], url, images[] Step 2: text, rating, date, reviewerName

Gotcha

Airbnb pricing varies by date. Always set checkIn and checkOut for accurate pricing. For market analysis, run multiple date ranges to capture seasonal variation.

Real estate lead scoring and agent routing

When: User wants to qualify inbound leads from listing portals by budget signals and urgency, then route them to the right agent.

Pipeline

  1. Search matching properties -> tri_angle/redfin-search
    • Key input: location, minPrice, maxPrice (from lead payload)
  2. Enrich lead with LinkedIn signals -> harvestapi/linkedin-profile-scraper
    • Key input: profileUrls (optional - use only when lead email resolves to a LinkedIn profile)

Output fields

Step 1: address, price, beds, baths, status, url Step 2: headline, currentCompany, experience[] (income/seniority signals)

Gotcha

The LinkedIn enrichment step is optional - only run it when the lead's identity is known and a LinkedIn profile URL is available. The core routing logic (hot/warm/cold tier + agent assignment) runs on the MLS webhook payload itself, with scraping used as enrichment. Lead scoring and routing output fields are AI-generated: leadTier, assignedAgent, routingReason.

Construction and pre-market property discovery

When: User wants to find new-construction projects or pre-market inventory before they appear on major listing portals.

Pipeline

  1. Scrape construction portals -> apify/playwright-scraper
    • Key input: startUrls (local MLS or construction project portal URLs), proxyConfiguration
  2. Extract clean text -> lukaskrivka/article-extractor-smart
    • Pipe: results[].url -> urls

Output fields

Step 1: raw HTML / structured page data Step 2: projectName, price, location, possessionDate, constructionStatus

Gotcha

No market-specific Actor exists for most construction portals (e.g., 99acres). Run apify actors search "real estate" --user-agent apify-agent-skills/apify-ultimate-scraper to check for community-built options before using apify/playwright-scraper. For JS-heavy portals, playwright-scraper is required. Step 2 cleans raw output into structured fields - pipe all Step 1 URLs through it.

Multi-source property comparison

When: User wants to compare listings across Zillow, Realtor, Zumper, and other US/UK sources.

Pipeline

  1. Aggregate listings -> tri_angle/real-estate-aggregator
    • Key input: location, propertyType, sources (Zillow, Realtor, Zumper, Apartments.com, Rightmove)

Output fields

Step 1: address, price, beds, baths, sqft, source, url, listingDate

Source: SKILL.md on GitHub

1 alert16d5 checks · Risk SAFE
  • Gen Agent Trust Hub16d

    The skill provides comprehensive instructions for orchestrating and interacting with web scrapers (Actors) via the Apify CLI. It provides structured playbooks for different data extraction workflows (B2B lead generation, brand monitoring, social media analytics, etc.) and contains standard guidelines on configuration, parameter passing, and error handling. No malicious behaviors, obfuscation techniques, or hidden actions were detected.

  • Socket16d

    No alerts

  • Snyk16d

    Risk: MEDIUM · 1 issue

  • Runlayer7mo

    2/2 files flagged

  • ZeroLeaks5mo

    Score: 93/100 · 2 sections analyzed

Signed by skilld at 97047d3. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub yesterday.

Activeupdated 3 months ago
  • SEO
  • apify
  • scraping
  • web-scraping
  • instagram
  • facebook
  • tiktok
  • youtube
  • linkedin
  • twitter
  • google-maps
  • reddit
  • lead-generation
  • competitor-analysis
  • brand-monitoring

README badge

README badge for apify/agent-skills/apify-ultimate-scraper

Scrapes data from 100+ Apify Actors covering Instagram, Facebook, TikTok, YouTube, LinkedIn, X/Twitter, Google Maps, Reddit, Airbnb, Yelp, and 15+ other platforms via the Apify CLI. Use for lead generation, competitor analysis, brand monitoring, influencer discovery, review analysis, or SEO intelligence on any public web platform.

Generated from the current SKILL.md.

Does this skill work with all platforms?
It covers ~100 Actors across 15+ platforms including Instagram, Facebook, TikTok, YouTube, LinkedIn, X/Twitter, Google Maps, Reddit, Airbnb, Yelp, and more. Not all platforms are equally supported; check the actor-index.md reference for your target platform.
What authentication is required?
You need an Apify account and CLI token. Authenticate via `apify login` (OAuth), set `APIFY_TOKEN` as an environment variable, or source from a .env file.
Do I need to install anything locally?
Yes, Apify CLI v1.5.0 or later must be installed via npm (`npm install -g apify-cli`).
Can I export results in different formats?
Yes. Results can be fetched as JSON or CSV using `apify datasets get-items` with the `--format` flag.
What should I read before running a scraper for the first time?
Check `references/actor-index.md` to find the right Actor for your platform, then read `references/gotchas.md` for common pitfalls specific to that Actor before running.

Generated from the current SKILL.md. These answers refresh after source changes.