@maschinenlesbar.org/openka-cli
    Preparing search index...

    Function extractPdfText

    • Extract the text layer of a PDF.

      Never throws for a merely difficult document: a page whose content stream will not decode contributes no text and one entry in problems. It throws only when the bytes are not a PDF at all, or when the document is encrypted — cases where there is nothing to be salvaged and the tier must abstain outright.

      Parameters

      • bytes: Buffer

      Returns PdfTextResult