API reference

Crawl

Crawl an entire site and return every page in the format you choose.

POSThttps://api.scrapeflow.dev/v1/web/crawl1 credit per page crawled

Starts an async crawl job. Poll the returned `job_id` or set a `webhook` to receive results when the crawl finishes. Runs on SQS-backed workers.

Request

curl -X POST ${SCRAPEFLOW_API_URL}/v1/web/crawl \
  -H "Authorization: Bearer $SCRAPEFLOW_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{ "url": "https://docs.stripe.com", "limit": 500 }'

Parameters

ParameterTypeDescription
urlrequiredstringThe root URL to crawl.
limitnumberMax pages to crawl. Default 100.
includePathsstring[]Only crawl paths matching these globs.
webhookstringURL to POST results to when the job completes.

Response

json
{ "success": true, "job_id": "job_a1b2c3", "status": "queued" }