Get the words without a document editor
The UTF-8 output includes a marker before each selected page, helping trace a passage to its source. The preview shows up to 10,000 characters; Copy result and the download include the complete extraction.
Extraction is not OCR
A photographed page may be readable to a person while having no text layer. If no selected page contains text, the tool stops with an explanation. In mixed documents, pages without text receive a clearly marked note rather than invented content.
A simple letter usually follows the expected line order. A two-column report, table or diagram may need manual rearrangement. Unusual font mappings can produce incorrect characters. Compare important passages with the source before using them.
Limits are 20 MB, 100 pages and one million extracted characters. The tool does not translate, summarize, fact-check or unlock documents. For a visual copy of a scan, try PDF to JPG.
Processing and sources
Inputs are processed on your device, without uploading or saving them to browser storage. Reset clears the result; downloaded files remain on your device. Website requests and permitted advertising or measurement are separate. See our privacy policy.
Technical reference: PDF.js page API.
Related tools
Merge PDF, split PDF, PDF to Word and PDF to JPG.