PDF.co API icon

PDF.co API

Generate PDF, extract data from PDF, split PDF, merge PDF, convert PDF. Fill PDF forms, add text and images to pdf and much more with pdf.co!

Make PDF Searchable or Unsearchable

AI-generated

Overview

This node processes PDF files to either make them searchable or unsearchable using OCR (Optical Character Recognition) technology. It is useful for converting scanned PDFs into searchable text documents or reversing that process to make PDFs unsearchable for privacy or security reasons. Practical examples include digitizing scanned documents for text search or protecting sensitive information by making text unsearchable.

Use Case Examples

  1. Making a scanned PDF searchable to enable text search and copy functionality.
  2. Making a PDF unsearchable to protect sensitive information from being copied or searched.

Properties

Name Meaning
Authentication Selects the authentication method to access the PDF processing service (API Key or OAuth2).
PDF URL The URL of the source PDF file to be processed.
Make PDF Searchable or Unsearchable Choose whether to make the PDF searchable or unsearchable.
OCR Language Name or ID Specifies the language for OCR when making the PDF searchable. Only applicable when making the PDF searchable.
Advanced Options Additional optional settings for the PDF processing such as custom output file name, page ranges, callback URL, output link expiration, password protection, HTTP authentication for source URL, and custom processing profiles.

Output

JSON

  • url - URL of the processed PDF file.
  • status - Status of the PDF processing operation.
  • pagesProcessed - Number of pages processed in the PDF.
  • fileName - Name of the output PDF file.

Dependencies

  • An API key credential or OAuth2 authentication for accessing the PDF processing service.

Troubleshooting

  • Ensure the PDF URL is accessible and correct to avoid download errors.
  • Verify that the authentication credentials (API key or OAuth2 token) are valid and have necessary permissions.
  • Check that the OCR language specified is supported and correctly set when making the PDF searchable.
  • If using advanced options like password or HTTP authentication, ensure the credentials are correct.
  • For callback URLs, ensure the endpoint is reachable and correctly configured to receive data.

Links

Discussion