Crawl4AI Plus Advanced icon

Crawl4AI Plus Advanced

Advanced web crawling and extraction with full Crawl4AI API control

SEO Metadata Extractor

AI-generated

Summary

Extract SEO metadata from a specified URL using customizable metadata types and extensive browser configuration options.

Inputs

  • URL (required) — The URL to extract SEO metadata from.
  • Metadata Types — Select which SEO metadata types to extract, including Basic Meta Tags, JSON-LD Structured Data, Language & Locale, Open Graph Tags, Robots & Indexing, and Twitter Cards.
  • Options — Options to include the original webpage text and/or raw HTML head section in the output.
  • Browser & Session — Configure browser engine (Chromium, Firefox, WebKit), JavaScript enablement, stealth mode, headers, proxy, cookies, user agent, viewport size, timeouts, headless mode, and many other advanced session and browser behaviors.
  • Crawl Settings — Control crawling behavior including cache mode, respecting robots.txt, CSS selector for content limitation, delay before returning HTML, exclusion of tags or external links, JavaScript code execution before extraction, wait conditions for page readiness, anti-bot features, retry attempts, and content word count thresholds.

Output shape

a JSON object containing the extracted SEO metadata and optionally including the original page text and raw HTML head section depending on options selected.

The output contains selected metadata types as structured data extracted from the page's SEO-related elements. The node uses a headless browser configured per input settings to load and parse the page, ensuring accurate metadata extraction even for dynamically generated content. All options to control browser and crawl behavior are supported, allowing tailoring to complex extraction scenarios.

Links

Discussion