Skip to content

PDF to Text – Extract Text from Any PDF

Get the words out of a PDF as plain text you can copy, edit or save as a .txt file. The PDF is read in your browser and never uploaded.

How to convert a PDF to text

  1. Add your PDF. Drop it onto the box above, or tap to choose a file.
  2. Pick the pages. Keep All pages, or tap Choose pages and type the pages you want, for example 1-3, 5 or 8-.
  3. Mark page breaks if you need them. Switch on Mark where each page starts to add a line like ----- Page 3 ----- before each page.
  4. Press Extract text. A progress bar shows which page is being read.
  5. Copy or save it. Use Copy to paste the text elsewhere, or Download .txt to save it as a UTF-8 text file.

Features

  • Readable output. Words are put back into lines, and a blank line is added where the spacing on the page shows a new paragraph and between pages.
  • Two-column pages are read column by column instead of mixing the two columns line by line. Headings that run across both columns stay in place.
  • Tables stay separated. Wide gaps inside a line become tabs, so rows paste cleanly into a spreadsheet.
  • Page ranges, including open-ended ranges such as 8- for "page 8 to the end".
  • Optional page markers, which you can switch on or off after extracting without reading the file again.
  • Counts at a glance. The number of characters, words and pages is shown above the text.
  • Scanned-page detection. Pages without a text layer are listed by number, so you know which ones need OCR.
  • Right-to-left support for Urdu, Arabic, Persian and Hebrew text.

Why use Fileora to get text from a PDF

Copying text straight out of a PDF viewer is fiddly. Selections jump between columns, line breaks land in the middle of sentences, and on a phone it's often impossible to select more than a paragraph. Extracting the whole document at once gives you clean text to quote, translate, search, summarise or reuse.

Fileora reads the PDF in your browser, so contracts, bank statements and exam papers never leave your device. There's no sign-up, no watermark and no page limit, and the tool works the same on a phone as on a laptop. Once you have the text, paste it into the word counter to check its length and reading time.

When the text doesn't come out

A PDF can look like text on screen but hold only pictures. Phone scans, faxes and photocopies saved as PDF are like this. When no page has any text, the tool says so; when only some pages are scans, it names them. To read scanned pages, convert them with PDF to JPG and then run the images through Image to Text (OCR).

If you are asked for a password, the PDF is encrypted. Open it with Unlock PDF using the password you know, then try again.

A few PDFs use fonts without a proper character map. Their text looks right on screen but extracts as odd symbols or not at all. There is no reliable way to recover those characters from the file itself, so OCR is the answer here too.

What to expect from the output

The result is plain text, so bold, italics, colours, links and images are not included. Headers, footers and page numbers are part of the page text, so they appear in the output as well. Words split with a hyphen at the end of a line stay split. Very complex layouts, such as magazine pages with text boxes and captions scattered around photos, may come out in an unexpected order, because the reading order is rebuilt from where the words sit on the page.

Frequently asked questions

How do I extract text from a PDF?

Add your PDF, keep All pages or pick Choose pages, and press Extract text. The text appears in a box below with Copy and Download .txt buttons.

Why does it say no text was found?

Your PDF is most likely a scan: each page is a picture of text, not text itself, so there is nothing to extract. Turn the pages into images with PDF to JPG, then read them with Image to Text (OCR). If only some pages are scans, the tool lists those pages and extracts the rest.

Can I extract text from only some pages?

Yes. Choose pages and type page numbers and ranges such as 1-3, 5 or 8-. A range with no end, like 8-, runs to the last page. Pages are extracted in the order you type them.

Does it keep the layout of the PDF?

It keeps lines, paragraphs and the reading order, but not fonts, colours, images or exact positions. Two-column pages are read one column after the other, and wide gaps inside a line, such as between table cells, become tabs.

Does it work with Urdu, Arabic and other right-to-left PDFs?

Yes, as long as the PDF contains real text. Right-to-left lines are put back in reading order from the position of each word, with English words and numbers kept in their own order. Check the result, because some PDFs store joined letters in a way that can come out swapped.

What if my PDF is password-protected?

A PDF that needs a password to open can't be read. Remove the password with Unlock PDF first, then extract the text. PDFs that open without a password but restrict copying are read normally.

Is my PDF uploaded to a server?

No. The text is extracted by your browser on your own device. Nothing is sent anywhere, so confidential documents stay private.

Can't find the tool you need, or something isn't working?

Tell us which tool you'd like next or what went wrong. We read every message, and requests decide what we build next.

Opens your email app. Please don't attach private files.