Slate PDF

Converting

How to Extract Text From a PDF (Free)

Extract text from a PDF for free — pull the readable text out of an electronic PDF into plain text or Markdown, and know what to do when your PDF is really an image.

HM

Hamza M.

PDF Workflow Editor

September 24, 20264 min readReviewed September 24, 2026

Extracting text from a PDF means pulling the words out of the file so you can copy, quote, reuse or search them outside the document. How to extract text from a PDF is, for an electronic PDF, a quick conversion rather than retyping — and for a scan it is a different problem entirely. This article covers both, and tells you honestly when the plain route will not work.

How to extract text from a PDF

For a PDF with a real text layer — one created from Word, a design tool, or a printed-to-PDF save — extraction is a single conversion. The tool reads the document and returns its words as re-usable text, in plain text or Markdown, with headings and lists intact. It runs entirely in your browser: no upload, no account, nothing leaves your device.

  1. 1Open the PDF to Markdown tool and add your PDF.
  2. 2Convert — the text layer is read and returned as Markdown (or plain text).
  3. 3Copy the output into a document, notes app, or AI tool and edit as needed.

Why the file has no text layer at all

The fastest way to see what you are dealing with: select a line of text with your cursor and press Ctrl+C. If a copy results, the file has a text layer and extracts cleanly. If the selection returns nothing, the page is an image — scanned, photographed or flattened on export. Copy-testing a document once tells you which pipeline to use, and it saves a failed conversion.

In the PDFWhat extracts
ParagraphsClean, run-together-per-line text with paragraph breaks
HeadingsHeading levels detected from size and weight
Bulleted and numbered listsReconstructed as real lists
Simple tablesStructured rows and columns when the grid is regular
Text baked into a scanNothing — no text layer exists until OCR runs

Plain text or Markdown — which should you extract?

Both start from the same text layer. Plain text is right when you are pasting into an email, a form or a quick note. Markdown is right when the text will live somewhere structured — documentation, a wiki, a static site — or when you are feeding a document to an AI tool, because headings and lists give the model structure to navigate by. If you only need a few lines, copying from the PDF directly is fastest and needs no conversion at all.

When your PDF is a scan

A scanned PDF holds no extractable text until it is run through OCR, which recognizes the letter shapes in the image and writes a searchable text layer into the file. That changes the document itself, and it is the honest route for scans rather than an "extract from image" hack. For one-off copying from a scan, the simplest path is the OCR tool that turns the page into readable, selectable text — the same pipeline the scanned-document posts below explain end to end.

Keep the original file safe

Extraction is read-only on its output — the source PDF is never altered. Still, export the result into your system and keep the original where it was. Re-extracting is free, so there is no reason to overwrite anything. If the document is confidential, the same privacy applies throughout: the conversion runs locally, and what you do with the text afterwards — including pasting it into a hosted service — is a separate decision.

Do it now

PDF to Markdown

Convert PDF content to Markdown text. It runs in this browser tab — your file is not uploaded anywhere.

Open PDF to Markdown

Frequently asked questions

Why can't I copy text from my PDF?+

The PDF is probably a scan or a flattened image with no text layer. Select a line and press Ctrl+C; if nothing comes back, run OCR to add a searchable, copyable text layer first.

Can I extract text from a scanned PDF?+

Not directly — the words are pixels in an image. OCR rebuilds a text layer so extraction works, which is the correct route for scans.

Does extracting text keep the formatting?+

Text yes, styling no. Headings and lists are reconstructed as structure, but precise fonts, spacing and page layout do not survive plain-text output — that is expected for this task.

Will tables and columns come out right?+

Simple tables extract as structured rows and columns. Two-column layouts flatten into one column, sometimes with lines interleaved, and merged-cell grids are the hardest case.

Is extracting text from a PDF free?+

Yes, the tool runs in your browser with no upload, no account and no charge — your document never leaves your device.

Keep reading

HM

Hamza M. · PDF Workflow Editor

Hamza has spent a decade working with print and digital documents — from prepress production to everyday PDF cleanup — and now writes practical, tested guides for Slate PDF.

Read more about how this site is run