Recognise text in scanned PDFs
The PDF OCR tool reads scanned PDF pages with Tesseract optical character recognition, entirely in your browser, and returns the recognised text as plain text files. Scan a contract, receive a photocopied form, archive an old letter — OCR makes their content searchable and copyable again.
Scanned PDFs are photographs of paper: you cannot select, search or quote a single word. OCR bridges that gap, recognising dozens of languages and outputting clean text with per-file organisation for multi-document batches.
Recognition runs on your device — important for confidential material such as personal correspondence, medical documents and financial papers, which are exactly the documents that tend to exist only on paper. Pair it with PDF to Text for born-digital files that already have a text layer.
Like every Piclizer tool, this one runs entirely in your browser — your files stay on your device.