About OCR PDF
OCR PDF turns a scan that is only pictures of pages, like a signed contract, an old manual or a stack of receipts, into a PDF you can search with Ctrl+F and copy text from. Every page goes through text recognition, and the words are placed as an invisible layer exactly over where they appear. The page you see doesn't change at all, with the same images, stamps and layout.
The engine reads English and Chinese. Once the text layer is there, Compare PDF and the word search in Redact PDF can work with the file too. Clear, straight scans give the best results, so run Deskew PDF first if the pages are crooked. The file is saved with _ocr added to its name.
How to use OCR PDF
- 1Upload the scanned PDF
Click Choose file or drop the scan here. Up to 50 MB free, 1 GB with VIP.
- 2Run OCR PDF
Click OCR PDF. Each page is recognized in turn, which takes a little while for long scans.
- 3Download and search
Save the file ending in _ocr and press Ctrl+F in any PDF reader to try it.
Why use Cubfile for this
- Layout untouched
The scanned page stays exactly as it was; only an invisible text layer is added.
- Search and copy
Find words with Ctrl+F, and select or copy text as in any digital PDF.
- English and Chinese
Printed English and Chinese are recognized, including mixed lines.
- Works with other tools
With a text layer in place, Compare PDF and Redact PDF's word search can use the file.