How to Extract Text from a Scanned PDF

A scanned PDF looks like a normal document, but it's actually just a picture of one -- which is why you can't select, copy, or search its text the way you can with a PDF someone typed. Here's how to tell the difference, and how to actually get usable text out of one.

What makes a PDF "scanned"

A PDF built from a document that was typed (a Word export, an invoice generated by software, a web page saved as PDF) carries an actual text layer: every character is stored as data, which is what lets you select, copy, and search it.

A scanned PDF is different. It's built by photographing or scanning a physical page and saving that image inside a PDF wrapper -- there's no text data at all, just pixels. To your PDF viewer, the whole page is one big picture, indistinguishable from a photo of a receipt.

How to check which one you have

The fastest test: open the PDF and try to select a sentence with your cursor. If a real selection highlights individual words, it has a text layer already -- you don't need OCR at all, just copy the text directly.

If nothing highlights, or your cursor just draws a selection box around the whole page like it's a photo, it's scanned -- which is exactly the case this guide, and OCR generally, is for.

Extracting the text

Scanquil's PDF to Text tool handles this directly: choose the PDF, and it reads every page on-device using OCR -- no upload, nothing sent to a server. Pages that already have a real text layer are read directly (instant, and more accurate than re-OCRing something that was never an image); only genuinely scanned pages go through the OCR pass.

Once it finishes, download the result as plain text. If you need the layout preserved too -- paragraphs, tables -- PDF to Word keeps that structure instead of flattening everything to plain text.

When a scan is genuinely hard to read

Free, on-device OCR handles most scans well, but a photocopy of a photocopy, a skewed phone photo, or faint carbon-copy print can trip it up. For exactly that case, PDF to Word (AI) reads each difficult page with a more capable model instead of the free engine, and gives you the result back as an editable Word document.

Frequently asked questions

Is my PDF uploaded anywhere?

No. The free PDF to Text tool reads your file entirely in your browser, on your own device -- it never leaves your computer.

Will this preserve the original formatting?

PDF to Text extracts plain text only, no layout. If you need paragraphs and tables preserved, use PDF to Word instead, which reconstructs the page's structure.

Is there a limit on how many pages?

The free tool handles PDFs up to 100 pages. Very large scanned PDFs will simply take longer, since each page needs its own OCR pass.

Does this work on mobile?

Yes -- it works in any modern browser on a phone, tablet, or desktop, with no app to install.

Related tools