Autonomous multi-page extraction into structured JSON. Use when the user wants website data matching a schema — pricing tiers, product listings — beyond a single-page scrape.
All-time #366Trending #2236Hot #6079First seen Mar 10, 2026
AI-powered autonomous extraction of structured data from complex multi-page websites.
Navigates sites intelligently to locate and extract data, returning results as JSON with optional schema validation Supports custom JSON schemas for predictable structured output, or freeform extraction when schema is not provided Offers two model tiers (spark-1-mini and spark-1-pro) with credit limits and optional waiting for inline results Best suited for multi-page extraction tasks; use simpler scrape or crawl skills for single-page or bulk extraction without AI reasoning
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
This skill is safe and provides a legitimate interface for using the Firecrawl CLI to perform autonomous data extraction from websites. All external resources and tools identified are associated with the skill's author and represent standard functionality.
AI-powered autonomous extraction. The agent navigates sites and extracts structured data (takes 2-5 minutes).
Quick start
# Extract structured data
firecrawl agent "extract all pricing tiers" --wait --json -o .firecrawl/pricing.json
# With a JSON schema for structured output
firecrawl agent "extract products" --schema '{"type":"object","properties":{"name":{"type":"string"},"price":{"type":"number"}}}' --wait --json -o .firecrawl/products.json
# Focus on specific pages
firecrawl agent "get feature list" --urls "<url>" --wait --json -o .firecrawl/features.json
Run firecrawl agent --help for the full option list.
Done when: the output file contains valid JSON answering the request — or a job ID was intentionally returned for later polling.
Job IDs
Omitting --wait returns a job ID. A UUID positional argument is auto-detected as a status check:
# Check once (equivalent to adding --status)
firecrawl agent "<job-id>"
# Wait on an existing job, polling every 10 seconds for up to 5 minutes
firecrawl agent "<job-id>" --wait --poll-interval 10 --timeout 300
# Cancel an active job
firecrawl agent "<job-id>" --cancel
Tips
Use --wait for inline results; omit it only when you want a job ID to poll later (see [Job IDs](#job-ids)).
Use --schema for predictable, structured output — otherwise the agent returns freeform data.
Agent runs consume more credits than simple scrapes. Use --max-credits to cap spending.
For simple single-page extraction, prefer scrape — it's faster and cheaper.