PDFTEQ Team
PDFTEQ Team Makers of PDFTeq's PDF tools
Published Updated
6 min read

The short answer

A scanned PDF is a picture of a page, so a voice reader has no text to speak. You need OCR (optical character recognition) to turn the picture into text first. PDFTeq's free PDF reader does that in your browser and then reads the result aloud. Expect a few seconds per page, and better results from clean, straight scans of printed text.

You open a PDF, press play in your reader, and nothing happens. Or you get an error about there being no text. It feels like the tool is broken, but the file is usually the problem. Somebody scanned a paper document, saved it as a PDF, and what you have is a stack of photographs of pages.

That's fixable, and you don't need paid software to do it. This guide covers how to confirm it's a scan, three ways to get it read aloud, and what to do when the result isn't perfect. For the general methods on phones and computers, see our guide to reading a PDF aloud on any device.

How do you know if a PDF is scanned?

Three quick checks. If all three fail, it's a scan.

  1. Try to select a sentence. Drag your cursor across some text. If it highlights word by word, the text is real. If it selects the whole page or nothing at all, it's an image.
  2. Search for a word you can see. Press Ctrl+F (Cmd+F on a Mac) and type it. No matches for a word that's clearly on the page means there's no text layer.
  3. Look at the quality. Slightly fuzzy letters, a tilted page, shadows near the edge, or a visible paper texture all point to a scan or a phone photo saved as PDF.

Some scanned PDFs have OCR text hidden behind the image. Those pass the checks and read aloud without any extra step.

Why can't voice readers read scans?

Text-to-speech works on characters, not pictures. When you scan a page, the scanner records what the page looks like. It doesn't know that a cluster of pixels is the letter "a". Tools like Adobe's Read Out Loud, Edge's Read aloud, and your phone's Speak Screen all need real text, so they either stay silent or skip the page.

OCR closes that gap. It analyses the picture, works out which characters are on it, and outputs text. Once you have text, any voice reader can speak it.

Method 1: How do you OCR and listen in your browser?

This is the quickest route if you don't have Acrobat Pro. With PDFTeq's online PDF reader:

  1. Drop your scanned PDF into the box.
  2. Wait while it checks each page. Pages with no selectable text are flagged, and you'll see an OCR prompt.
  3. Choose the document's language. Getting this right matters more than anything else.
  4. Press Run OCR. A progress bar shows which page it's on, and you can stop at any time.
  5. When it's done, pick a voice and press play. Click any sentence to jump to it.

What to know: everything runs in your browser, so your PDF isn't uploaded. The OCR engine and its language data are downloaded when you start. Each run handles up to 30 scanned pages, so split longer files into parts and do them one at a time. It doesn't save a searchable PDF or an audio file; it reads the text aloud and that's it.

Method 2: Can Acrobat read a scanned PDF?

Only after OCR. The free Acrobat Reader can't recognise text in a scan. Acrobat Pro can, through its Scan & OCR tool. Once the file has a text layer, View > Read Out Loud works normally.

This makes sense if you already pay for Acrobat Pro, or if you also want a searchable PDF you can keep. If you just want to listen to one document, it's more setup than it's worth.

Method 3: What about your phone's camera tools?

If you have the paper original, you can skip the PDF entirely. Google Lens (in the Google app on Android and iPhone) can recognise text through the camera and has a Listen option that reads it aloud. It works page by page, which is fine for a letter and painful for a 40-page report.

How do you get better OCR results?

OCR is only as good as the scan. The same tool can read one file almost perfectly and mangle another, and the difference is nearly always the image. A few things help:

  • Rescan at a higher resolution. Around 300 dpi is a common target for text documents.
  • Keep the page straight and flat. Tilt, curved book spines and shadows all cause errors.
  • Use good contrast. Dark text on a clean, light background reads best. Faded or coloured backgrounds are harder.
  • Choose the right language. Reading English text with a Hindi model, or the other way round, produces nonsense.
  • Don't expect handwriting to work. OCR engines like this are built for printed text.
  • Listen critically to numbers and tables. That's where mistakes hide. For contracts, medical papers or anything important, check the original.

What about Gujarati, Hindi and other languages?

Two separate things have to work, and people often only think about one.

  • The OCR language decides how well the page is recognised. PDFTeq offers English, Hindi and Gujarati, along with English plus Hindi or Gujarati for mixed documents. Recognition for Indian scripts tends to be less accurate than for English, so expect more mistakes.
  • The voice decides whether the text is spoken properly. Voices come from your device, so if there's no Gujarati or Hindi voice installed, the text may be read in the wrong accent or skipped. Add the voice in your phone or computer's speech settings first.

Comparison of the options

We make PDFTeq, so we're biased. That's why the table lists its limits as well.

Method Cost Best for Main downside
PDFTeq (browser OCR) Free, no signup A quick listen to a scanned document 30 pages per run, device voices, no saved file
Acrobat Pro + Read Out Loud Paid Also needing a searchable PDF Cost and setup
Google Lens Listen Free A few pages of paper One page at a time

Frequently asked questions

Yes, but only after OCR turns the page pictures into text. A voice reader can't read a scan directly. PDFTeq offers OCR in your browser when it finds pages with no selectable text.

The most common reason is that it's a scan: an image of a page with no text layer. Try selecting a sentence or searching for a word. If neither works, the file needs OCR first.

Not by itself. The free Reader can't run OCR on a scanned file, so Read Out Loud has nothing to speak. Acrobat Pro can recognise the text first, after which Read Out Loud works.

No. OCR runs in your browser. The OCR engine and its language data are downloaded when you start OCR, and voices marked "online" in the voice list may send the spoken text to your browser's vendor.

On a laptop, a few seconds per page. Phones are slower. PDFTeq processes up to 30 scanned pages per run, so split longer files into parts.

Usually the scan itself: low resolution, skewed or shadowed pages, small print, handwriting, or the wrong language setting. Rescanning at a higher resolution and picking the correct language fixes most problems.

PDFTeq offers Gujarati and Hindi OCR, but recognition tends to be less accurate than for English, so expect more errors. You also need a Gujarati or Hindi voice installed on your device to hear the result read properly.

No. It recognises the text so it can read it aloud, but it doesn't save a new searchable PDF or an audio file.

Got a scanned PDF?

Drop it in and run OCR in your browser. No signup, and the file isn't uploaded.

Read it aloud
PDFTEQ Team

Written by the PDFTEQ Team

We build and maintain PDFTeq's browser-based PDF tools, including the OCR and read-aloud features described here. OCR quality varies a lot with scan quality, so treat any result as a draft and check anything important against the original.

More guides