About this tool
A scanned page is just a picture — you cannot search it or copy from it. OCR reads the characters in the image and adds an invisible text layer behind the page, so the document looks identical but becomes searchable.
How to use this tool
- Select a scanned PDF or an image.
- Choose the language of the document — accuracy depends heavily on this.
- Click Run OCR. Recognition takes a few seconds per page.
- Download the searchable PDF, or copy the extracted plain text.
How it works
Recognition runs locally using a WebAssembly build of the Tesseract engine. The language data downloads once and is cached by your browser, so subsequent runs start faster.