If you upload a scanned PDF here, it gets rejected. There is no OCR on this platform, and a scan is not text — it's a picture of a page. A voice generator can only read characters it can find in the file, and a scanned page has none, no matter how sharp the image looks on screen.
That is the limitation, stated plainly, and it is not going to change by trying a different file name or a smaller file. What follows is how to check whether your PDF is actually a scan, how to fix it in the tools you probably already have, and what to watch for once you've fixed it and are ready to generate the audio.
Is Your PDF Actually a Scan?
You can find out in about five seconds without opening any special software. Open the PDF in any reader and try to click-and-drag to select a line of text. If a text cursor appears and a sentence highlights, the file has a real text layer and you're fine. If nothing highlights, or your cursor turns into a little rectangle for drawing a selection box over the image, it's a scan.
A second check that catches the same thing: press Ctrl+F (or Cmd+F) and search for a word you can see printed on the page. A text-based PDF finds it instantly. A scanned PDF finds nothing, because there is nothing to search — just pixels.
Heads up: Some PDFs are a mix — a cover page that's a scanned signature sheet, followed by pages of real text. The whole file is only readable to us where the text layer exists. Check a page from the middle and the end, not just the first page.
Fixing a Scanned PDF
The fix is OCR, and you almost certainly already have access to it somewhere. OCR — optical character recognition — reads the shapes of letters in the scanned image and turns them into actual, selectable text. We don't do this step for you, but three places usually already do.
Your scanner's own export
Most office scanners and scanning apps have a "searchable PDF" or "PDF with OCR" export option sitting right next to the plain "PDF" option. If you scanned the document yourself, re-export with that option checked.
A phone scanning app
The built-in scanning tools on modern phones (in the camera app or notes app) run OCR automatically when you scan a document and save it as a PDF. Re-photographing a printed page this way is often the fastest fix.
Your operating system's text recognition
Recent versions of desktop operating systems include a built-in "recognize text" or "extract text" feature for images and PDFs. It's usually a menu option in the default photo or PDF viewer, not a separate app you need to install.
Once the OCR pass has run, save the result as TXT or DOCX rather than re-saving as a PDF. A plain text or Word file guarantees the text layer travels cleanly and strips out any leftover image data you don't need.
Two Things OCR Reliably Ruins
OCR is good at recognizing letters and bad at understanding page layout, so it hands you back two specific problems almost every time. Fix both before you generate audio, not after — they're a five-minute find-and-replace job now and an annoying re-listen later.
Turn the fixed document into audio
Upload the OCR'd TXT or DOCX file and generate a spoken version in a voice that suits the material.
Open the text-to-audiobook toolWhat Formats Actually Work
Once a document has a real text layer, four formats are accepted for turning it into audio: EPUB, TXT, PDF (text-based, not scanned), and DOCX. There's no MOBI support and no audiobook file (M4B) output — you get MP3, and WAV if you're on Pro or above.
Textbooks and readings
Usually already text-based unless they've been photocopied and re-scanned at some point.
Government and legal forms
Often scanned from a paper original, which is why they fail the select-a-line test more often than most files.
Old e-books
EPUB files from years ago are almost always text-based already and rarely need OCR at all.
Printed reports
A report that was scanned for archiving rather than authored digitally is the classic case that needs fixing first.
What Converting a Document Costs
Generation is priced in points, and points work out to roughly one point per second of finished audio. The exact cost is calculated from character count rather than duration, so treat the per-second figure as a useful estimate rather than an exact rule — a page with a lot of long words will cost a few more points than the seconds of audio suggest, and a page of short, punchy sentences will cost a few less.
That difference is expected, not a bug. It's also a reason to fix the OCR problems above before you generate: a document still cluttered with running headers and hyphen breaks costs you points reading text nobody wanted narrated in the first place.
Once the Text Is Clean
With a real text layer in hand, you have the full range of document-to-audio tools available. A textbook chapter, a scanned government form you converted to DOCX, or an old EPUB you've had sitting around all go through the same path from here.
Pick a voice that fits how the document will actually be used. A dense technical manual benefits from a clear, unhurried reader; a report someone will listen to on a commute can carry a bit more warmth. The six voices below cover both ends of that range — press play on any of them before you commit.
| Voice | Accent | Tone | Rating |
|---|---|---|---|
| Warm Natural Female Explainer | Neutral American | Warm, natural | ★★★★★ |
| Articulate Male Explainer Voice | Neutral American | Articulate, insightful | ★★★★★ |
| Polished Approachable British Female | British | Polished, approachable | ★★★★★ |
| Lucid Male Explainer Voice | Neutral American | Clear, measured | ★★★★★ |
| Calm British Male Narrator | British | Calm, cordial | ★★★★★ |
| Professional Clear American Male | Neutral American | Clear, professional | ★★★★★ |
More Document-to-Audio Tools
PDF to audiobook
Convert a text-based PDF into narrated audio, chapter by chapter.
Accessible PDF to audio
Give a text-based PDF a spoken version to support readers who prefer listening.
DOCX to audiobook
Turn a Word document straight into audio without reformatting it first.
TXT to audiobook
The most reliable path for anything you've just run through OCR.
EPUB to audiobook
Narrate an e-book file with its chapter breaks intact.
Frequently Asked Questions
Can I upload a scanned PDF directly?▾
No. Scanned PDFs are rejected on upload because they contain an image of the page rather than actual text, and there's no OCR built into the platform to extract it for you.
How do I know if my PDF is a scan or real text?▾
Try to select a line of text with your cursor, or search the document for a word you can see. If nothing highlights and the search finds nothing, it's a scan and needs OCR first.
Where do I get OCR done?▾
Check your scanner's export settings for a "searchable PDF" option, use your phone's built-in document scanning feature, or use your operating system's text recognition tool. Save the result as TXT or DOCX.
Will the audio read out page numbers and headers?▾
It will if you leave them in. OCR typically pulls in running headers, footers, and page numbers as regular text, so remove them from the document before generating audio.
What file types can I upload once the text is clean?▾
EPUB, TXT, text-based PDF, and DOCX. There's no MOBI support, and the output is MP3 (WAV on Pro and above) — not an M4B audiobook file.
Get your document ready and hear it read aloud
Once your PDF has a real text layer, upload it — or the TXT or DOCX you converted it to — and generate a clean narrated audio file.









