Merge PDF
Combine several PDFs into one file in the order you choose, without leaving your browser.
Extract the text from a PDF, page by page, into a box you can copy. The PDF is read in your browser.
Updated
Your browser is preparing the tool. It runs 100% locally.
Upload a PDF and the tool reads its text layer page by page and puts the result in a box you can copy. It works on PDFs that contain real, selectable text. A scanned PDF is just an image of the page with no text layer, so it returns nothing and needs OCR instead. Your browser opens and reads the file, so it never goes to a server.
This tool pulls the words out of a PDF so you can reuse them. It reads the document's text page by page and gives you plain text, ready to copy into a note, an editor or a search box.
Use it on PDFs that already have selectable text, such as reports, exported documents and articles, to get the text out without retyping it.
The tool opens the PDF with a rendering engine and, for each page, reads the text items the PDF stores (the actual characters placed on the page). It joins them, puts a blank line between pages, and shows the result in a copyable box.
This works because a normal PDF carries a text layer: the words are stored as text, even though the page looks like a fixed image. A scan or photograph of a page has no such layer, so there is nothing to read and the result is empty.
for each page: read the stored text items → join them with spaces
pages separated by a blank line · a scanned page with no text layer → nothinga 3-page text PDF → its text, page by page, in a copyable boxIt reads the existing text layer only. A scanned PDF (an image of text) has no text layer, so it returns nothing and needs OCR. Columns and tables run together, because items are joined in reading order.
The PDF stays on your device. Your browser opens it and reads the text, with no upload and no copy kept anywhere.
There is no account, and the page saves nothing. Refresh or close the tab and both the PDF and the extracted text are cleared from memory, so it is fine for private documents.
For a PDF with a real text layer the extraction is fast and faithful: the words come out as stored, page by page. That covers most exported and generated PDFs.
A scanned PDF returns nothing. If the page is an image of text, there is no text layer to read and the box stays empty. Those documents need an OCR tool, which reads the picture of the words.
The layout isn't kept. Because the text items are joined into one stream, multi-column pages, tables and side-by-side text can come out interleaved or run together, even though the words themselves are right.
Spacing can vary. The tool joins text items with spaces, which usually reads well, but some PDFs split or space words oddly, which can add or drop a space compared with the original.
Copy text out of a report or document.
Pull a section of text to quote elsewhere.
Get the text out so you can search or analyse it.
Take PDF text into a word processor or your notes.
For a scanned PDF, use an OCR tool, which reads the image and produces text. To keep the layout, export to a format that preserves it. This tool reads only the text layer the PDF already has.
PDF.js reads the text layer the PDF already has. A scanned PDF is only a picture of text with no text layer, so it comes back empty and needs OCR instead.
Combine several PDFs into one file in the order you choose, without leaving your browser.
Turn plain text into a multi-page PDF set in Helvetica on US Letter pages.
Copy the pages you name, like 1-3,5, out of a PDF into a new file, in your browser.
Turn every page of a PDF into its own JPG image, rendered in your browser with PDF.js.
Turn JPGs or PNGs into one PDF, one image per page, each page sized to its image, with no re-compression.
Re-save a PDF with object-stream compression. Text-heavy files get smaller and nothing on the page changes.
Your file is processed on this device and never uploaded, and there's no account to create.