๐Ÿ“„ PDF ยท Updated October 11, 2026 ยท 8 min read

How to Turn a Scanned PDF Into Text With Free OCR

Scanned PDF Editable text ๐Ÿ”ค

To turn a scanned PDF into text, run it through OCR, because a scan is a picture of a page with no real characters to copy. GrabCast's OCR PDF does that in your browser: choose the document's language (up to three), optionally set a page range, drop in the file, and download a searchable PDF that looks exactly like the scan, or copy all the text and save it as a .txt. The Tesseract.js engine reads each page on your device, so a lease, a tax notice or a medical record never leaves your computer. This guide covers how to confirm you have a scan, what the settings do, how to get the cleanest result, and the limits worth knowing before you rely on the text.

๐Ÿ”ค Try OCR PDF now โ€” freeOpen โ†’
Scanned PDF turned into text with OCR PDF: recognized letter text beside the page preview with highlighted words
A scanned letter turned into editable, searchable text, all in the browser.
๐Ÿ’ก Why a scanned PDF will not give up its text

A scanned PDF is a photograph of each page wrapped in a PDF container. Your viewer sees a grid of pixels, not words, which is why you cannot highlight a sentence, search for a name or paste a clause into an email. Ordinary PDF-to-text converters return nothing, because there is nothing to extract. Optical character recognition fixes that by studying the shape of every letter and rebuilding the words as real characters. The payoff is large: a 40-page scanned contract becomes searchable in minutes, figures from a printed statement can go into a spreadsheet, and the text can finally go into a translator or an AI assistant that rejects image-only files.

How to tell whether your PDF is scanned

Confirm what you have first, because the right tool depends on it.

Some PDFs are mixed, with typed pages plus a few scanned signature or appendix pages. OCR PDF handles that case automatically, as the next section explains.

Turn a scanned PDF into text with OCR PDF

The tool opens the PDF with pdf.js, reads each page with Tesseract.js and rebuilds the file with pdf-lib, all inside your browser. The engine and language data download once; your document does not.

The searchable PDF keeps every page exactly as scanned and adds an invisible text layer placed over the matching words, so you can search, select and copy while the page looks unchanged.

Get the cleanest possible OCR result

Clean scans at 200 to 300 DPI usually come out nearly word-perfect. Faint, blurry, handwritten or sideways pages read worse. A little preparation costs less than fixing garbled text line by line.

If a single page comes out badly, rescan just that page rather than fighting the whole document.

Limits, and what to check before you rely on the text

OCR text is a very good draft, not a certified copy of the original. Plan a short proofreading pass, weighted toward the parts that carry meaning.

A spellchecker pass flags most stray errors in seconds. Then compare totals, dates, names and account numbers against the page image before the text goes into anything important.

Step-by-step

1234
1Confirm the PDF is a scan by trying to select a line of text in your viewer.
OCR PDF settings: English chosen as the document language, the Pages box left empty for all pages, and the drop zone below
Pick the language of the document (up to three) and leave Pages empty to read every page.
2Open OCR PDF, choose up to three languages used in the document and, if you like, a page range such as 1-3, 5, 8-end.
OCR finished for scanned-letter-ocr-ocr.pdf: 1 page made searchable, 169 words, 95% avg. confidence, with a Download searchable PDF button
The scan is read in your browser; the summary shows words found and the average confidence.
3Drop in the scanned PDF and wait while each page is read on your device, watching the progress and time left.
Recognized text of the 1998 Harbor Point letter about repairing the sea wall, with Copy all text and .txt buttons
Proofread the recognized text, especially numbers, dates and names, then copy it or save a .txt file.
4Download the searchable PDF, or Copy all text or save the .txt, then proofread numbers, dates and names.
Scanned page 1 with every recognized word boxed in purple, showing what the OCR engine read
The page preview boxes each recognized word, so gaps show where OCR missed something.

Common mistakes to avoid

โš ๏ธRunning a scan through a plain PDF-to-text converter and concluding the file has no text worth saving.
โš ๏ธUsing a blurry, low-resolution copy of a copy and expecting word-perfect output.
โš ๏ธReplacing a digitally signed original with the OCR version, which invalidates the signature.
โš ๏ธForcing OCR on pages that already have text, then finding every word duplicated when you search.

Pro tips

โœ“Scan at 300 DPI in black and white or grayscale for text documents; color adds size without helping recognition.
โœ“OCR only the pages you need with a page range when a long file has one important section.
โœ“Keep both files: the original scan as the record and the searchable PDF as the working copy.
โœ“After OCR, compress the searchable PDF if it is too big to email; scans are usually the heaviest part.
โœ“Run a spellcheck on the .txt to surface most recognition slips in one pass.

Frequently asked questions

Is my scanned PDF uploaded to a server?

No. It is opened with pdf.js, read by Tesseract.js and rebuilt with pdf-lib inside your browser. Only the engine and language data download the first time.

Will the searchable PDF look different from my scan?

No. The original pages are kept as they are, with an invisible text layer added over the matching words. Because the file is re-saved, existing digital signatures become invalid.

Which languages can OCR PDF read?

Twenty, including English, Spanish, French, German, Italian, Portuguese, Russian, Ukrainian, Arabic, Hindi, Chinese, Japanese, Korean and Vietnamese. You can combine up to three in one document.

Can it read handwriting?

No. It is built for printed and typed text. Handwritten notes and signatures are left as they are, while the printed text on the same page is recognized.

Is there a page or file size limit?

There is no fixed limit and it is free with no sign-up, but everything runs in your browser's memory. Use a page range to process a very long document in parts.

๐Ÿ“Œ Bottom line

A scanned PDF is a stack of pictures until OCR turns it into text. Check that you have a scan, pick the right languages in OCR PDF, feed it a clean, upright file, and download a searchable PDF or a .txt. Then proofread the numbers and keep the original as your record.

Open OCR PDF โ†’

Related guides

Browse more: all PDF guides ยท OCR PDF