PDF to Text Extractor
Get selectable text from a PDF without uploading it. Choose your pages, check the result against the document, and copy or download plain text.
1 Open your PDF
Or drop one PDF here · up to 20 MB and 200 pages. Passwords stay in memory and are cleared after entry.
No PDF loaded.
Choose a PDF to begin.
2 Choose pages & text layout
Page ranges are read in document order. Layout is approximate: columns, tables and unusual fonts may need manual corrections.
0 pages selected
3 Review & download
Show original PDF page
It may be blank, scanned or use unsupported text encoding. This extractor does not perform OCR.
For a scan, convert the PDF page to an image, then use Image to Text / OCR. Check any recognized text before using it.
Exports contain complete extracted text, including pages beyond the preview limit. Canceled or failed runs retain completed pages and mark their downloads as partial.
Using PDF to Text Extractor
Extract the selectable text layer from a PDF into plain text. Choose pages, inspect the result and optionally compare it with a rendered source page.
- Choose or drop one PDF up to 20 MB and 200 pages. Enter its password if prompted.
- Select page checkboxes, all/odd/even pages or a strict range such as 1, 3-5. Choose detected line breaks or continuous text per page, then extract.
- Review each completed page. Copy text or download combined TXT, separate page text in a ZIP or a JSON report. Pages without selectable text show links to PDF to Images and OCR.
Example
A 12-page report with selectable text: enter 1, 3-5 and apply the range. Extract pages 1, 3, 4 and 5 in document order. Download a combined TXT with page headings, or a ZIP containing four page TXT files.
Questions & answers
Is my PDF uploaded or saved?
No. Parsing, extraction, previews and ZIP packaging run in this browser using locally hosted scripts. The PDF, passwords and text are not stored in browser storage. Clear file & text removes the current document and results. Downloads contain the extracted text.
Can this extract text from scanned PDFs?
It extracts selectable text already present in a PDF, including an existing OCR text layer. It does not recognize pixels. A page with no detected text may be blank, scanned or use unsupported encoding; convert it to an image and use Image to Text / OCR.
Will it preserve columns, tables and fonts?
Output is plain text. Detected line breaks and approximate spacing can help readability, but reading order, columns, tables and unusual fonts may be imperfect. Compare important passages with the source preview and correct the text before using it.
Does it support different languages?
It extracts Unicode text where the PDF contains usable character mappings. There is no language selector or OCR model here. Some fonts, missing mappings, right-to-left layouts and vertical writing can give incomplete or incorrectly ordered output.
What does Continuous text per page do?
It replaces detected line breaks with spaces within each page. It does not reconstruct paragraphs, remove hyphens or understand the document. Combined downloads can optionally include page headings.
What are the limits?
One PDF up to 20 MB and 200 pages, with up to 20000 text fragments and 100000 UTF-16 text units on each page and 1000000 units overall. The preview shows up to 60000 units per page; copy and downloads contain the complete extracted page. Loading times out after 30 seconds, excluding time waiting for a password, and extraction after 90 seconds.
Can I unlock a password-protected PDF?
Enter a password you know when asked. Incorrect passwords allow up to three attempts before loading is canceled. The password field is cleared after entry and passwords are not saved.
What happens if extraction is canceled or fails?
Completed pages remain available. The result summary and filenames mark the extraction as partial, and the JSON/ZIP report lists requested and completed pages. Changing page selection or text layout clears old results so they cannot be confused with a new extraction.
Guides for this tool
Examples and checks to help with your next step.
Help improve this tool
Report a problem or suggest an improvement
Describe the issue without pasting private tool input. Feedback goes to our admin inbox.
