How do I extract text from a PDF?
To extract text from a PDF, drop it into FixDoks's PDF to Text tool, which reads the stored text page by page and rebuilds it into readable lines. Turn on page separators if needed, then click Copy text or Download .txt. A scanned PDF with no text layer needs OCR PDF first. This runs entirely in your browser.
PDF to Text at a glance
| Input | One PDF file |
|---|---|
| Output | Plain text, or a .txt file |
| Processing | In your browser; your files are not uploaded |
| Price | Free, no watermark |
| Download | Free Google sign-in |
| Works on | Chrome, Edge, Safari and Firefox on Windows, Mac, Android and iPhone |
How to use PDF to Text
- Drop your PDF onto the box.
- FixDoks extracts the text page by page and rebuilds the lines for you.
- Turn on page separators if you want a marker between each page's text.
- Click Copy text, or Download .txt to save the file.
How does text extraction work?
A PDF with real text, meaning it was created from a word processor, a website, or any software that draws actual letters rather than a picture of letters, stores that text internally along with its position on the page. FixDoks reads through the file page by page, pulls out that stored text, and rebuilds it into readable lines and paragraphs in the order it appears, so you get plain text you can copy, search or reuse without retyping anything.
Page separators
Turning on page separators inserts a clear marker between each page's text in the output, which is useful when you need to know where a page ends, for example when quoting a specific page number from a long report, or when re-splitting the text later. Leave it off for a continuous flow of text, which reads more naturally if you are pasting it into another document.
Copying vs downloading
Copy text puts everything on the clipboard in one click, ready to paste into an email, a document or a search box. Download .txt saves it as a plain text file, which is useful for archiving, for feeding into another tool, or when the extracted text is too long to comfortably paste by hand. The live character count next to the output helps you judge length, for example checking against a submission form's character limit.
Scanned PDFs have no text to extract
A PDF made by scanning a paper document, or by photographing pages with a phone, is really just a picture of a page saved as a PDF: there is no text layer stored inside it at all, only pixels. FixDoks checks for this and will tell you plainly when a page has no extractable text, rather than returning blank or garbled output. In that case, the text needs to be recognised from the image first using optical character recognition; use OCR PDF to turn a scanned PDF into one with a real, searchable text layer, then run this tool again.
What can affect the result
- Multi-column layouts, like newspapers or academic papers, can sometimes extract in an unexpected reading order, since the text is stored by position on the page rather than by visual column.
- Tables often lose their grid structure, coming out as text with gaps or line breaks rather than neat columns.
- Headers, footers and page numbers are extracted just like body text, so long documents may include repeated header or footer text on every page.
Typical uses
| Goal | Why this helps |
|---|---|
| Quoting a paragraph elsewhere | Copy exact text instead of retyping it. |
| Searching a long PDF | Paste the text into a text editor and use its search function. |
| Feeding text into another tool | Download the .txt and open it wherever plain text is needed. |
If your goal is a Word document you can keep editing with the original formatting, use PDF to Word instead, since this tool intentionally produces plain text without fonts, colours or layout.