8 min read

Extract Text from Multiple PDFs at Once — Free

Bulk extract text from PDFs and images: one .txt per file, page markers kept, OCR for scans. Pull embedded images in bulk too. Free, no upload.

Quick answer

To extract text from multiple PDFs at once, open ihatepdf.cv/extract-text and select or drop all the PDFs together. Turn on OCR mode if any are scans, then press Extract. You get one .txt file per PDF, with page markers, in a single ZIP. Extract Images works the same way for embedded pictures.

Getting the words out of a pile of PDFs is the first step of a lot of work: building a research corpus, loading documents into a search tool or an AI assistant, reviewing a disclosure bundle, moving old content to a new website, or simply grepping a year of invoices for one supplier. Copying and pasting from each file is slow and loses track of which text came from where.

The Extract Text tool processes a whole batch and gives you one plain text file per document, with page markers, in a single ZIP. It reads scans and photos too, with OCR. And for the pictures inside the PDFs, Extract Images does the same in bulk. Nothing is uploaded by either.

How to extract text from many files at once

  1. Open ihatepdf.cv/extract-text.
  2. Select or drop all the files together — PDFs, and images such as JPG, PNG, WebP, TIFF, GIF or BMP. With more than one file the tool switches to batch mode.
  3. Turn on OCR mode if any of the PDFs are scans. Leave it off for PDFs with real text; it is much faster.
  4. Press "Extract N files".
  5. Download all: one ZIP with a name_extracted.txt per file.

What the text files look like

Each PDF becomes one UTF-8 text file. The text of every page follows a marker such as --- Page 3 ---, so you always know where a passage came from, and you can cite or return to the right page of the original. Images produce a single block of text.

The text comes out in the order it is stored in the PDF, which for ordinary documents is reading order. On complex layouts — two-column papers, sidebars, pages with many text boxes — lines from different columns can interleave. If you need the structure preserved, convert to Word instead (PDF to Word in bulk rebuilds columns, tables and headings).

Scans and photos in the same batch

You do not need to separate digital PDFs from scans and photos:

If what you want from your scans is searchable PDFs rather than text files, batch OCR keeps each page as it looks and adds the text on top.

Extracting images in bulk

Extract Images pulls out the photos, logos and figures embedded in each PDF — the original image data, not a screenshot of the page. In a batch:

A small number of PDFs store images in JPEG 2000 or JBIG2, two rare formats browsers cannot decode; those images are reported as skipped. More on how extraction works: extracting images from a PDF at full quality.

Choosing between text, Word and searchable PDF

Working with the output

Plain text is the most portable format there is. The ZIP's files can be searched in one go with your operating system's search, a code editor's "find in files", or grep on the command line; imported into a spreadsheet; or loaded into a notes app or a document assistant. Because each file keeps its document's name and page markers, any line you find leads straight back to its source.

Frequently asked questions

How do I extract text from multiple PDF files at once?

Open ihatepdf.cv/extract-text, select or drop all the PDFs together and press Extract. Each PDF becomes its own .txt file and you download them all as one ZIP.

Can it extract text from scanned PDFs?

Yes. Turn on OCR mode and every page is read with OCR. Images such as JPG and PNG are always read with OCR.

Does the text keep page numbers?

Yes. Each page's text follows a "--- Page N ---" marker, so you can trace every passage back to its page.

Can I extract all the images from multiple PDFs?

Yes, with Extract Images. Each PDF's embedded images are saved at their original quality, in a folder per PDF, with small decorative images skipped.

Why is the text from a two-column PDF mixed up?

Plain text follows the order the PDF stores it, which on multi-column layouts can interleave lines. Convert to Word to keep columns and structure.

Are my files uploaded?

No. Text and image extraction both run in your browser, and the files never leave your device.

→ Use these tools

Extract Text → Extract Images → OCR PDF → PDF to Word →

Try all tools free — no sign-up, no watermark

40+ free PDF tools. Files never leave your device.

Open ihatepdf →

Related guides

Batch Process PDFs

All 21 tools with batch mode.

Extract Text from One PDF

The single-file guide.

Extract Images from a PDF

Full-quality image extraction.

Batch OCR PDFs

Searchable PDFs instead of text files.