Dump the existing text layer from a digital PDF in this tab. If you cannot select words in Acrobat or the browser, it is probably a scan — use English OCR on a page image, not this extractor.
You need a quote out of a spec or contract and you can already highlight the words. That is this page: copy the text layer in the current tab, without a convert-to-Word upload.
It is not OCR (pixels). It is not PDF to JPG (pictures of pages).
Scan vs digital
Open the PDF. If the cursor can select a sentence, extract here. If you can only rubber-band a bitmap, it is a scan — even if the filename says invoice.pdf.
Permissions can also block copy in some readers while a library still sees the layer. If extraction returns empty, treat it as a scan or a locked file.
How to use it
- Drop the digital PDF.
- Extract in this tab. Skim for hyphenation and broken columns.
- If you needed one page as an image for chat, export that page.
Typical ceiling is about 32 MB.
Honest limits
- No Excel tables, no layout fidelity.
- Compress-PDF’s JPEG rebuild will not be extractable afterward.
- Not legal advice about whether you may copy the text.
To keep a subset of pages as a PDF, split first, then extract.