OCR for scanned PDFs
Turn a scan into text
No database: AIMRAN doesn't store your files, text or results
What is OCR for scanned PDFs?
A scanned PDF is really a photo of each page: you can't search for words or copy the text. OCR reads those images and recognizes the letters to turn them into real text.
You get the recognized text to copy or download and, if you like, a searchable PDF: the same document, looking the same, but one you can now search with Ctrl+F and select text in.
How to use it
- Drag in the scanned PDF.
- Choose the document's language: Spanish, English or both.
- Turn on “Create searchable PDF” if you want the PDF with text.
- Click “Recognize text” and wait for all the pages to be processed.
- Copy or download the text and, if you asked for it, the searchable PDF.
Advantages
- Turns scans into searchable PDFs without changing how they look.
- Recognition in Spanish, English or both.
- Processes every page of the document in one go.
- Free and no sign-up.
Technical details
Recognition uses tesseract.js with the LSTM engine, automatic page segmentation and preserved spaces between words. Each page is rendered with pdf.js at about 300 DPI, with a maximum of 4200 px per side, which is the resolution recommended for Tesseract. For the searchable PDF, Tesseract generates an invisible text layer per page that is overlaid with pdf-lib on the original page, without touching the image and respecting its rotation. Language data is downloaded the first time (about 2 MB for Spanish, 3 MB for English and 5 MB for both). Accuracy depends on scan quality; handwriting is not reliably recognized.
Frequently asked questions
How do I convert a scanned PDF into editable text?
Drag in the PDF, choose the language and click “Recognize text”. Then copy the text or download it as .txt to edit it.
What is a searchable PDF?
It is a PDF that looks just like the original scan but carries the recognized text hidden underneath, so you can search for words and copy passages.
How do I know if my PDF needs OCR?
If you can't select the text with your mouse or searching finds nothing, the PDF is scanned and needs OCR.
How do I get better results?
Scan at a good resolution, with straight, well-lit pages, and choose the document's correct language.
More PDF tools
- Merge PDF · Combine several PDFs and reorder pages by dragging
- Split PDF · By ranges, every page or every N pages
- Compress PDF · Three compression levels with adjustable quality
- Rotate pages · Rotate specific pages or all of them
- Delete pages · Remove the pages you don’t want
- Reorder pages · Drag the thumbnails to change the order
- Extract pages · Pull the selected pages into a new PDF
- PDF to JPG/PNG · One image per page, with your choice of DPI