pdfextract API
    Preparing search index...

    Interface TextPage

    Text and diagnostics in top-left page points, with CropBox, UserUnit and rotation applied.

    interface TextPage {
        blocks: TextBlock[];
        fullText: string;
        height: number;
        ocrStatus: "disabled" | "not-needed" | "performed" | "partial" | "failed";
        pageNumber: number;
        rotation: 0 | 90 | 180 | 270;
        warnings: ExtractionWarning[];
        width: number;
    }
    Index
    blocks: TextBlock[]

    Geometrically ordered text blocks; complex semantic reading order is not guaranteed.

    fullText: string

    Text derived from ordered blocks on this page.

    height: number

    Canonical page height in points after crop, user-unit scaling and rotation.

    ocrStatus: "disabled" | "not-needed" | "performed" | "partial" | "failed"

    Coverage outcome. A failed region cannot be hidden by another region succeeding.

    pageNumber: number

    Original one-based PDF page number.

    rotation: 0 | 90 | 180 | 270

    Original PDF page rotation in clockwise degrees.

    warnings: ExtractionWarning[]

    Page-local limitations and collected errors.

    width: number

    Canonical page width in points after crop, user-unit scaling and rotation.