PDF to text: extract, copy and save

Read selectable PDF text and save a plain text file with page breaks kept. Files never leave your device: this job runs in your browser. Nothing is uploaded.

This tool does not perform OCR. A scanned PDF needs an existing text layer. Columns, tables and reading order may need manual correction. Hidden text in the PDF can appear in the result, so review it before sharing.

Limit: one unencrypted PDF, 100 MB and 200 pages; up to 10 million output characters.

Choose a PDF to begin.

Text extraction uses PDF.js by Mozilla (Apache-2.0).

How to use this tool

Choose an unencrypted PDF and select Extract text. Review the result, then copy it or save a UTF-8 .txt file. Page boundaries are kept as form-feed characters, including pages with no text.

Frequently asked questions

Is PDF to text free, and do you upload my file?

It is free without an account. Extraction, copying and the text download happen in your browser. Files never leave your device; nothing is uploaded.

Can I extract text from a scanned PDF?

Only if the PDF already has a text layer. This tool does not perform OCR. If no text layer is found, the page says so. Pages containing only scans, pictures or blank space are reported individually when other pages contain text.

Are page breaks and formatting kept?

Page boundaries are saved as form-feed characters in the UTF-8 text file. Text lines and spaces are inferred from the PDF’s text items. Fonts, images and exact page layout are not included. Columns, tables, unusual encodings and reading order may need manual correction.

Why can extracted text include hidden content?

PDF text extraction reads the text layer, which can include text outside a crop or behind shapes. Visually hiding text does not remove it. Review the result before sharing; use proper PDF redaction to permanently remove sensitive information.

What are the limits?

Choose one unencrypted PDF up to 100 MB and 200 pages. Text output is limited to 10 million characters. Password-protected, damaged or oversized files are rejected; partial results are not saved. Large PDFs may exceed your browser’s memory. Clipboard access may require HTTPS or your browser’s Copy command.