All skills
firecrawl avatar

/firecrawl-agent

@4d145f7 official
by firecrawlfirecrawl/cli643 stars
110

Autonomously navigate websites and extract structured data across pages. Use when the task requires navigation or no suitable ready-made workflow or data provider covers it.

Use this Skill: https://skilld.dev/gh/firecrawl/cli/firecrawl-agent

This session only. Nothing lands on disk.

SKILL.md

≈48 tokens always: the name and description. ≈849 when used: this file.

firecrawl agent

AI-powered autonomous extraction. The agent navigates sites and extracts structured data (takes 2-5 minutes).

Before starting autonomous extraction for structured records or listings, check firecrawl search alexandria '<data you need>' for a ready-made workflow or data provider. Inspect a matching contract with firecrawl list <provider> <capability> --pretty and execute with firecrawl scrape --alexandria <provider>/<capability> --options '<input JSON>' if it covers the task. Use the exact provider, capability, and input fields from that contract. Continue with Agent when no suitable tool exists or the task requires autonomous navigation.

Quick start

# Extract structured data
firecrawl agent "extract all pricing tiers" --wait --json -o .firecrawl/pricing.json

# With a JSON schema for structured output
firecrawl agent "extract products" --schema '{"type":"object","properties":{"name":{"type":"string"},"price":{"type":"number"}}}' --wait --json -o .firecrawl/products.json

# Focus on specific pages
firecrawl agent "get feature list" --urls "<url>" --wait --json -o .firecrawl/features.json

Run firecrawl agent --help for the full option list.

Done when: the output file contains valid JSON answering the request — or a job ID was intentionally returned for later polling.

Job IDs

Omitting --wait returns a job ID. A UUID positional argument is auto-detected as a status check:

# Check once (equivalent to adding --status)
firecrawl agent "<job-id>"

# Wait on an existing job, polling every 10 seconds for up to 5 minutes
firecrawl agent "<job-id>" --wait --poll-interval 10 --timeout 300

# Cancel an active job
firecrawl agent "<job-id>" --cancel

Tips

  • Use --wait for inline results; omit it only when you want a job ID to poll later (see Job IDs).
  • Use --schema for predictable, structured output — otherwise the agent returns freeform data.
  • Agent runs consume more credits than simple scrapes. Use --max-credits to cap spending.
  • For simple single-page extraction, prefer scrape — it's faster and cheaper.

See also

Alexandria session feedback

To report an Alexandria session outcome or a provider/capability gap, use firecrawl alexandria feedback --rating good|partial|bad --url <website> --requested-functionality '<what was needed>' --objective '<the underlying goal of the task>' --rationale '<what happened>' --json. Use observed results in the rationale. No job ID is needed; this session feedback has no job-age deadline and no credit refund. Optional --provider-feedback and --capability-feedback JSON arrays describe specific gaps; inspect firecrawl alexandria feedback --help for their fields. Use the capability issue missing_capability when a provider exists but lacks the needed capability, and new_capability_request (with requestedFunctionality) to ask for one.

Source: SKILL.md on GitHub

1 warningtoday5 checks · Risk SAFE
  • Gen Agent Trust Hubtoday

    The firecrawl-agent skill allows for autonomous website navigation and data extraction. While it processes content from external websites, which is a common vector for indirect prompt injection, this behavior is central to its purpose and the skill uses the vendor's own supported tools.

  • Sockettoday

    No alerts

  • Snyktoday

    Risk: MEDIUM · 1 issue

  • Runlayer6mo

    1 file scanned · No issues

  • ZeroLeaks5mo

    Score: 93/100 · 2 sections analyzed

Signed by skilld at 4d145f7. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub 19 hours ago.

Activeupdated yesterday
What it can do
Runs commands
All 2 allowed tools
Bash(firecrawl *)Bash(npx firecrawl-cli *)
  • CLI
  • firecrawl
  • web-scraping
  • data-extraction
  • json
  • structured-data
  • autonomous-agent

README badge

README badge for firecrawl/cli/firecrawl-agent

Extracts structured JSON data from complex multi-page websites using AI navigation, handling pricing tiers, product listings, and directory entries. Useful when manual scraping would require navigating many pages or when you need the agent to infer where data lives on a site.

Generated from the current SKILL.md.

What's the difference between firecrawl-agent and firecrawl-scrape?
firecrawl-agent uses AI to autonomously navigate multi-page sites and extract structured data, taking 2-5 minutes. firecrawl-scrape is a simpler, faster, and cheaper single-page extraction tool.
Does this support custom output schemas?
Yes. Pass a JSON schema via --schema or --schema-file to get predictable structured output. Without a schema, the agent returns freeform data.
Which models are available?
Two options: spark-1-mini and spark-1-pro, specified with the --model flag.
How do I control costs?
Use --max-credits to cap the credit spend for a single agent run, since agent extraction consumes more credits than simple scraping.
Do I need to wait for results?
Use --wait to block and get results inline. Without it, the command returns a job ID and you must poll separately.

Generated from the current SKILL.md. These answers refresh after source changes.