NextCloud PDF icon

NextCloud PDF

Lit et remplit les champs de formulaire (AcroForm) de PDFs stockés sur Nextcloud

Get Text

AI-generated

Summary

Extracts the raw text content of a specified PDF file stored on Nextcloud, returning detailed text broken down page by page.

Inputs

  • Depuis (required) — Choose how to specify the PDF file: either from a selectable list of files in a Nextcloud folder or by providing the file path as an expression.
  • Dossier — When selecting from a list, pick the folder to filter available PDF files. Defaults to root '/'.
  • Fichier PDF — When selecting from a list, choose the specific PDF file within the selected folder.
  • Chemin du fichier PDF — When specifying by path, provide the full path to the PDF file on Nextcloud.

Output shape

a single JSON object containing the PDF path, total page count, the full extracted text, and an array of page-specific text entries.

The returned JSON includes 'pdfPath' (the path used), 'text' (all text concatenated), 'pageCount' (total number of pages), and 'pages' (array of objects with 'page' number and corresponding 'text'). Each item in the workflow output corresponds to one processed input item, with robust error handling if the PDF cannot be processed.

Examples

Example 1: Extract text from a specific PDF by selecting it from the root folder

Set 'Depuis' to 'Depuis une liste', choose '/' as 'Dossier', and select the desired 'Fichier PDF'. The node returns the text content page by page.

Example 2: Extract text by directly specifying PDF file path

Set 'Depuis' to 'Par chemin (expression)' and provide the full PDF path in 'Chemin du fichier PDF'. The node fetches and extracts its text content.

Links

Discussion