OCR PDF
Make scanned PDFs searchable by adding a text layer. 100% processed in your browser.
The language model downloads once (~10 MB) on first use.
Drop a scanned PDF here
or click to browse
or click to browse
Frequently Asked Questions
Is my file uploaded to a server?
No. The OCR engine (Tesseract.js) runs entirely inside your browser using WebAssembly. Your PDF never leaves your computer.
Why is the first run slow?
The English language model (~10 MB) is downloaded on first use and cached for future sessions.
Does this work on handwritten text?
No. Tesseract is designed for printed text. Handwriting accuracy is very low.
Does OCR change what my page looks like?
No. Your page stays exactly as it was -- same logo, same layout, same page count. Only the recognized text itself is replaced in place with a clean, selectable copy; anything that isn't text (images, borders, decorative graphics) is never touched.