Can't copy text from a PDF? Why and what to do
Run a ten-second check to learn whether your PDF holds real text or only images, then pick the fix that matches, with your file staying in the browser.
A PDF refuses to give up its text for one of three reasons: the pages are scanned images, the author set restrictions, or the fonts don't map back to letters when you copy. If the PDF has real text, you can extract the text from the PDF with your own file, see it with a character and word count, and copy it or save it as a .txt. If the pages are scans, the tool tells you instead of handing back an empty result. The file is processed in your browser and isn't uploaded.
First, find out which case you have.
The 10-second check
Open the PDF and press Ctrl+F (Cmd+F on a Mac). Search for a word you can see on screen.
- Found it? The document has real text. The problem is the viewer, a restriction, or the order of what you paste.
- Nothing found? You're looking at a picture of text: a scan or a screenshot saved as PDF. No viewer will copy it.
That one check saves you from trying fixes that don't apply.
Scanned pages: you need OCR
A PDF made by a scanner or from photos of a page stores each page as an image, so there are no letters to select. The fix is OCR, which recognizes text inside the image. The Extract Text tool here does not do OCR: it detects which pages are images and lists them so you know their content isn't in the result. That is a plain limitation. The guide to extracting text from a scanned PDF walks through the OCR route.
Restrictions set by the author
A PDF can carry permissions that block copying or printing. If it's your own document and you know the password, remove the PDF password and try again. If it belongs to someone else and the restriction is deliberate, respect it: the author chose not to allow reproduction, and asking for an editable version is the right route.
If the PDF needs a password to open, the Extract Text tool takes it in an optional field and uses it only in your browser.
You can copy, but it comes out wrong
Two flavors here. In the first, the text arrives in the wrong order, with columns interleaved or a line break in the middle of every sentence. A PDF stores positioned fragments of text, not paragraphs, so what you copy follows how the page was drawn. We cover that, plus the Markdown fix for pasting into a chatbot, in why PDFs paste badly into ChatGPT. In the second, you get strange symbols or nonsense letters. That happens when an embedded font lacks the mapping from its glyphs to actual characters. There's no reliable fix from the text layer, and sometimes the only way out is to treat that page as an image and run OCR.
How to extract the text, step by step
- Open Extract Text from PDF and drop your file (up to 50 MB).
- If the PDF needs a password to open, type it in the optional field.
- Click Extract text. Past 300 pages, the tool warns you that it may take a while and asks you to confirm.
- Check the character and word counts, and the notice about pages with no text if one appears.
- Copy to the clipboard or download the
.txt.
The output is plain text with no formatting, so headings, tables and bold are gone. To keep basic structure, use PDF to Markdown, which is built to keep headings and lists.
Getting the text into Word or Docs
Once you have the .txt, open it in any editor and paste it into Word or Google Docs. Expect to redo the layout, because plain text carries no fonts, columns or images. For a long document it still beats retyping, and the word count shown after extraction gives you a quick sense of how much text you are moving.
Need a page as a picture instead?
For a snapshot of the page, text included, use PDF to JPG. It won't give you text, but it gives you an image you can feed to an OCR. If you want the photos embedded in the document, see how to get images out of a PDF.
Frequently asked questions
Why can't I select text in my PDF?
Usually because the pages are images from a scan, because the author applied restrictions, or because the viewer blocks it. A quick test is searching a word with Ctrl+F: if it's found, the text exists and the cause is something else.
How do I copy text from a scanned PDF?
You need OCR, which reads the letters inside the image. The Extract Text tool here doesn't do OCR: it flags the image-only pages so you know their text isn't in the result.
Can I extract text from a password-protected PDF?
Yes, if you know the opening password: enter it in the optional field and it's used only in your browser. It isn't a way around restrictions on someone else's document.
Why does copied text have odd line breaks and mixed columns?
A PDF stores positioned fragments of text, not paragraphs. When you copy, the order depends on how the page was drawn. Converting to Markdown usually produces a cleaner result.
Does the extracted text keep its formatting?
No. It comes out as plain text without headings, bold or tables. For basic structure, use the Markdown conversion.
Is my PDF uploaded to a server?
No. Text is extracted in your browser and the file never leaves your computer. You can check the Network tab in developer tools: analytics requests show up, but no POST or PUT the size of your file.
Pull the text out of your PDF
Get all the text at once, with a word count, to copy or download as .txt. Free, no signup.
Open Extract Text from PDF →Related tools
- Extract Text from PDF: all the text at once, to copy or download as .txt.
- PDF to Markdown: keeps headings and lists, ideal for pasting into an AI chat.
- Unlock PDF: remove the password from a document of yours that you know.
- PDF to JPG: turn pages into images, for example to run them through OCR.
You might also like: open a PDF without Adobe and what data a PDF can hide.