How to OCR a PDF (Make It Searchable)

A step-by-step guide to making scanned PDFs searchable with OCR: pick document languages, text layer added locally in your browser, files never uploaded.

Updated 2026-09-14·2 min read

OCR your PDF now

A scanned PDF is a photograph of text: it looks readable, but as far as every search box, copy-paste, and text extractor is concerned, the pages are empty. OCR (optical character recognition) fixes that by recognizing the letters and adding an invisible text layer on top of the scan. This guide shows you how with AmiPDF's OCR PDF tool, which runs the entire recognition in your browser: the file never leaves your device, which matters for contracts, medical records, and ID documents. You pick one or more document languages, the tool recognizes each page, and the output is the same PDF, now searchable and copyable.

Running OCR, step by step

1

Open the OCR tool and add your scan

Go to the AmiPDF OCR PDF tool and drop your scanned PDF onto the upload area. Scanned documents up to the tool's page limit are accepted; if your scan is a long book, cut it into parts first with split PDF and process each chunk.

OCR PDF – Open the OCR tool and add your scan
2

Pick the document languages

Select every language that appears in the document, up to the per-run limit. Choosing two or three languages when the text genuinely mixes them improves recognition; adding languages the document doesn't contain only slows the run. The engine and language packs download once on first use and are cached afterwards.

OCR PDF – Pick the document languages
3

Run the recognition

Click start and watch per-page progress. Recognition happens on your device with WebAssembly; a 10-page scan typically takes a few minutes depending on your machine. Keep the tab open until the run completes; the tool assembles the output automatically.

OCR PDF – Run the recognition
4

Download the searchable PDF

Save the result; it's the same visual document, now with an invisible text layer underneath. Search works in any PDF reader, copy-paste returns real text, and you can pull the content out later with PDF to Markdown. A button also offers the raw recognized text as a .txt file for quick checking.

OCR PDF – Download the searchable PDF

Getting the best results

Clean scans recognize better

Straight, evenly lit, 300 DPI pages are the sweet spot. Skewed phone photos lose accuracy at the edges; if the source is a phone camera, straighten and crop first; the crop PDF tool can trim dark borders that confuse recognition.

Check the text preview

OCR is statistical: 100% accuracy isn't guaranteed, especially with unusual fonts or low contrast. Use the extracted-text view to spot-check a few paragraphs before relying on search, and re-run with adjusted languages if the output looks garbled.

Skip OCR for text-based PDFs

If the tool tells you the document already has a text layer, you don't need it; extraction tools like PDF to Word work directly and finish in seconds.

Frequently asked questions

Is my file really not uploaded anywhere?

Correct — recognition runs in your browser via WebAssembly. The only network traffic is the one-time download of the OCR engine and language packs, cached for future runs.

How many languages can I pick?

Up to the per-run limit shown on the page. Mix languages freely, for example English and Chinese in one bilingual contract, but skip languages that aren't in the document; they only slow the run down.

Will OCR change how my scan looks?

No. The pages keep their original appearance; the recognized text is added as an invisible layer positioned under the scan. You see the same document, but search and copy now work.

Make the scan searchable

Local OCR, your files stay on your device. Free, no signup.

OCR your PDF now