Skip to main content

Search OfficeIMO

Enter a topic, API type, or PowerShell command.

Documentation Search all OfficeIMO/

API Reference

Class

PdfImageOcrExtensions

Namespace OfficeIMO.Pdf.Ocr
Assembly OfficeIMO.Pdf.Ocr
Modifiers static

Recognizes standalone raster images through the canonical positioned-text reconstruction pipeline.

Inheritance

  • Object
  • PdfImageOcrExtensions

Remarks

Recognition uses the source resolution, rather than rendering a PDF intermediate. The resulting logical document can be passed to existing Word, Excel, HTML and other format adapters. Animated and multi-page images are rejected; supply separate static sources when every frame or page is required.

Methods

public static async Task<PdfSearchableOcrResult> MakeSearchableAsync(PdfImageDocumentSource image, IOcrEngine engine, PdfOcrMergeOptions options = null, CancellationToken cancellationToken = default) #
Returns: Task<PdfSearchableOcrResult>

Creates an in-memory searchable PDF preserving the image and adding all accepted OCR words.

Parameters

image OfficeIMO.Pdf.PdfImageDocumentSource requiredposition: 0
engine OfficeIMO.Ocr.IOcrEngine requiredposition: 1
options OfficeIMO.Pdf.Ocr.PdfOcrMergeOptions = null optionalposition: 2
cancellationToken System.Threading.CancellationToken = default optionalposition: 3
public static Task<PdfSearchableOcrReview> PrepareSearchableOcrAsync(PdfImageDocumentSource image, IOcrEngine engine, PdfOcrMergeOptions options = null, CancellationToken cancellationToken = default) #
Returns: Task<PdfSearchableOcrReview>

Recognizes an image and retains a private image-page snapshot for review and searchable PDF generation.

Parameters

image OfficeIMO.Pdf.PdfImageDocumentSource requiredposition: 0
engine OfficeIMO.Ocr.IOcrEngine requiredposition: 1
options OfficeIMO.Pdf.Ocr.PdfOcrMergeOptions = null optionalposition: 2
cancellationToken System.Threading.CancellationToken = default optionalposition: 3
public static async Task<PdfOcrMergeResult> ReadWithOcrAsync(PdfImageDocumentSource image, IOcrEngine engine, PdfOcrMergeOptions options = null, CancellationToken cancellationToken = default) #
Returns: Task<PdfOcrMergeResult>

Recognizes a captured image and reconstructs editable text and tables without writing an output file.

Parameters

image OfficeIMO.Pdf.PdfImageDocumentSource requiredposition: 0
engine OfficeIMO.Ocr.IOcrEngine requiredposition: 1
options OfficeIMO.Pdf.Ocr.PdfOcrMergeOptions = null optionalposition: 2
cancellationToken System.Threading.CancellationToken = default optionalposition: 3