If every text-to-speech tool you try says your PDF has no text, it’s probably a scan. The file holds pictures of the pages and no actual words, so there’s nothing for a voice to read.
A text PDF stores characters: this is an “a”, in this font, at this spot on the page. A scanned PDF stores a photo of each page, wrapped up so it opens like a document. To you they look the same. Text to speech reads characters, so a scan gives it nothing to work with.
Some scans already have a hidden text layer added, often called a “searchable PDF”. Those will read aloud, recognition errors and all.
If it’s an old book, a clean digital text may already exist. Project Gutenberg has tens of thousands of proofread public-domain books, and they’ll always read better than text recognised from a scan. Aloud searches Gutenberg and Open Library directly. More on the free catalogue.
OCR (optical character recognition) turns the page images back into text.
Live Text can pick up text in a photo. Open a photo of the page, press and hold on the text, choose Select All, then Copy, and paste it into Notes or Pages. It’s accurate and it runs on the phone, but it’s far too slow for a whole book. For a letter or a handful of pages it’s ideal.
OCR text is only as good as the scan. Read a page or two of the new PDF before you commit a long trip to it, and look for:
If the first pages are poor, rescanning flatter and brighter usually helps more than any OCR setting.
Import the recognised PDF like any other. If OCR didn’t take, Aloud reports the file as a probable scan instead of opening a silent, empty book, so you’ll know straight away.
After OCR, Aloud reads the PDF in chapters with your place saved, narrated on your iPhone. Your first import is free.
Optical character recognition. It's software that looks at a picture of text, works out the letters and stores them as real text you can select.
No. If a PDF has no selectable text, Aloud tells you it's probably a scan instead of importing an empty book. Run OCR on it first and import the result.
Rarely. Clean modern print comes out well. Old typefaces, faint photocopies, curved pages and handwriting cause errors, and the voice reads those errors exactly as written.
It means uploading the document to someone else's server. That hardly matters for a public-domain book. For contracts, medical letters or anything private, use an offline tool on your own computer.