PDF.co API icon

PDF.co API

Generate PDF, extract data from PDF, split PDF, merge PDF, convert PDF. Fill PDF forms, add text and images to pdf and much more with pdf.co!

Convert From PDF

AI-generated

Overview

This node converts PDF files from a given URL into various formats such as CSV, HTML, JPG, JSON (AI-powered or simple), PNG, Text, TIFF, WEBP, XLS, XLSX, and XML. It supports advanced options for customization including page selection, OCR language, extraction region, output file naming, and callback URLs for asynchronous processing. This node is useful for automating PDF data extraction and format conversion workflows, such as extracting tables from invoices, converting reports to editable formats, or generating images from PDF pages.

Use Case Examples

  1. Convert a PDF invoice from a URL to CSV format to extract tabular data for accounting.
  2. Convert a PDF report to text for content analysis and indexing.
  3. Convert scanned PDF documents to searchable text using OCR with specified language.
  4. Convert PDF pages to JPG images for use in presentations or web content.

Properties

Name Meaning
Authentication Method of authentication to access the PDF conversion service (API Key or OAuth2).
PDF URL The URL of the PDF file to convert.
Convert Type The target format to convert the PDF into, such as CSV, HTML, JPG, JSON, PNG, Text, TIFF, WEBP, XLS, XLSX, or XML.
Advanced Options Additional options to customize the PDF conversion process, including file name, pages to process, OCR language, extraction region, callback URL, output expiration, inline response, line grouping, unwrapping lines, HTTP credentials for source URL, and custom processing profiles.

Output

JSON

  • data - The converted output data from the PDF in the selected format.
  • url - The URL to the converted file if output is provided via a link.
  • pagesProcessed - The pages of the PDF that were processed during conversion.

Dependencies

  • An API key or OAuth2 credentials for authentication with the PDF conversion service.

Troubleshooting

  • Ensure the PDF URL is accessible and correct; HTTP authentication credentials may be required if the URL is protected.
  • Verify that the selected convert type is supported and correctly specified.
  • If using OCR, ensure the correct language is selected to improve accuracy.
  • Check callback URL and webhook configurations if using asynchronous processing to receive output data.
  • Custom profiles must be valid JSON strings; invalid JSON will cause errors.

Links

Discussion