PDF OCR

Recognize scans and export a ZIP containing TXT plus a searchable PDF text layer.

Runs on the server· Files: 1 · 50 MB
Files

or choose files from this device. Processing starts only when you press the button.

Server
01 Choose files02 Review03 Process04 Download

Drop files here

or choose files from this device. Processing starts only when you press the button.

Options
Saved presets0
Choose files

How it works

PDF OCR re-checks the output before download. OCR is approximate and runs on the server with Tesseract. Scans beat born-digital text extraction.

  1. 01Add a supported PDF file. The real type is checked before PDF OCR runs, not only the extension.
  2. 02Use the controls that belong to this task: OCR language(s). Files over the stated limits are rejected.
  3. 03Check size, format and page or pixel counts on the result before downloading. Compression tasks keep the original when they do not save space.

Limits and support

Up to 1 file, 50 MB per input, 300 pages. Runs in the browser where the codec is available; the output is checked before download.

Privacy note

This task uses a server engine. You must consent before upload. Results expire after 15 minutes.

Output

PDF OCR re-checks the output before download. OCR is approximate and runs on the server with Tesseract. Scans beat born-digital text extraction.

Options

PDF OCR controls: OCR language(s). Limits: 1 file(s), 50 MB each, 300 pages.

Common questions

Which files can I use?

PDF OCR accepts PDF and creates ZIP. The actual file type is checked before processing. OCR is approximate and runs on the server with Tesseract. Scans beat born-digital text extraction.

Does this task upload my file?

It uses server-side processing. The workspace asks for explicit upload consent before sending the file. This task uses a server engine. You must consent before upload. Results expire after 15 minutes.

Which options are available?

Available controls include: OCR language(s). Limits: 1 file(s), 50 MB each, 300 pages.