ScrapegraphAI icon

ScrapegraphAI

Turn any webpage into usable data with the ScrapeGraphAI v2 API — scrape, extract, search, crawl, monitor, history, credits.

Actions19

Crawl → Get Status

AI-generated

Overview

This node operation retrieves the status of a specific crawl job from the ScrapegraphAI service. It is useful for monitoring the progress or current state of a web crawling task that has been initiated. For example, users can check if a crawl job is still running, completed, or encountered errors.

Use Case Examples

  1. Check the status of a crawl job by providing its unique crawl ID to monitor its progress.
  2. Use the simplified response option to get a concise summary of the crawl status instead of the full raw data.

Properties

Name Meaning
Crawl Job The unique identifier of the crawl job to retrieve the status for. It can be selected from a list of existing crawls or provided directly as a UUID string.
Simplify Whether to return a simplified version of the crawl status response instead of the raw data. Defaults to true.

Output

JSON

  • status - The current status of the crawl job, such as running, completed, or failed.
  • crawlId - The unique identifier of the crawl job.
  • details - Additional details about the crawl job status, which may include progress metrics or error messages.

Dependencies

  • Requires an API key credential for ScrapegraphAI to authenticate requests.

Troubleshooting

  • Ensure the provided Crawl Job ID is a valid UUID; otherwise, the node will throw a validation error.
  • If the crawl job does not exist or has been deleted, the node may return an error indicating the job was not found.
  • Network or authentication issues may cause the node to fail; verify API credentials and network connectivity.

Discussion