PDF to speech

Import a text-based PDF, clean it up, and listen to it or download it as audio.

PDFs are built for printing, not for listening. Reports, papers, manuals and course readers come with page numbers, running headers, footnotes and multi-column layouts that sound odd when read in order. And some PDFs are just scanned images with no text in them at all. This page explains how to import a PDF, what to clean up before generating, how to tell a scanned PDF from a text-based one, and how to turn a long document into audio you can listen to on the go.

Open the tool → Opens the tool with MP3 as the download format and highlights the Import text button so you can load your PDF straight away.

Importing the PDF

Drag the PDF onto the text box, or click Import text and choose the file. The text is extracted on your device and appears in the box, where you can read and edit it before anything is spoken. Nothing is uploaded, which matters for contracts, medical letters, coursework and other private documents.

If you only need part of the document, delete the rest from the box before generating. It is often quicker to import the whole file and trim it than to try to copy a few pages out of a PDF viewer.

Text-based versus scanned PDFs

A quick test: open the PDF in any viewer and try to select a sentence with your mouse. If you can highlight individual words, the PDF has real text and will import. If the whole page highlights as one block, or nothing highlights, it is a scanned image. Scanned PDFs have no text to read, so the box will come up empty or nearly empty.

For a scanned document, you need to run it through a text recognition step in another program first, then import the resulting file or paste the text. The tool does not do text recognition itself.

Cleaning up PDF text for listening

PDF text often carries leftovers from the page layout. Look for repeated headers and footers such as the document title or "Page 12 of 40" and delete them, since they will otherwise be read aloud every page. Footnote numbers stuck to the ends of words and hyphenated words split across lines are also worth a quick pass.

Two-column layouts, such as academic papers and newsletters, sometimes come out with lines interleaved from both columns. Skim the imported text; if paragraphs look jumbled, fix those sections before generating. Tables and figure captions rarely make sense as speech, so it is usually better to remove them or replace them with a short sentence summarizing the point.

Long reports and papers

There is no character limit, so a long report can go in as one piece. Playback starts while later parts are still being generated. For very long documents, consider splitting by chapter or section and downloading each as its own MP3, which makes it easier to pick up where you left off on a phone. See long text to speech for more on this.

For technical documents, try a speed of 0.9 so figures and terms have time to register. Numbers, percentages, dates and common units like km and MB are read in their spoken form automatically, which helps a lot with reports. Acronyms and product names may need respelling in the box if they come out wrong.

Listen now or save as audio

You can press play and listen in the page, click any part in the parts list to jump to it, or download the result as MP3 for a commute or walk. If you would rather follow along with text on screen, SRT subtitles are also available. For Word documents, the same approach applies; see Word to speech.

Frequently asked questions

Why is the text box empty after importing my PDF?

The PDF is most likely a scan, meaning each page is an image with no text layer. The tool can only read PDFs that contain real text.

Is my PDF uploaded to a server?

No. The file is read on your device and the text is never sent anywhere. After the first use, the tool also works offline.

Is there a page limit?

No. There is no character limit and no daily quota. Longer documents just take longer to generate, and slower devices and most phones take longer than laptops.

Can it read PDFs in other languages?

Only English. A French or German PDF will import fine, but the words are read with English pronunciation, which makes them hard to follow — the voices are American and British English only.

Can I convert a whole folder of PDFs at once?

No. Batch processing is not available; import and generate one document at a time.

Which voice should I use for a PDF?

Heart is the default and the most natural-sounding. For long study sessions, try a few voices with the short listen button in the voice menu and pick one you can listen to for a while.