Extract text from a PDF,free, in your browser
Drop a PDF and get a plain .txt file. You see the extracted text first, and the document never leaves your device.
The PDF is read inside your browser tab. It is not uploaded, so there is no copy of your document anywhere but your own device.Free to use: one PDF at a time, up to 50 MB, any number of pages, and no signup.
Drop your PDF
Drag and drop a PDF or click to browse.
Extract
Every page is read in turn and the text is joined in reading order.
Preview and download
Check the text looks right, then save it as a .txt file.
How to extract text from a PDF
Extraction is quick when the PDF holds real text. The sections below cover the two questions that bring most people here: why a scan gives nothing back, and why the layout does not survive.
Extract the text from a PDF
Drop one PDF onto the upload card, or click it to browse. Files can be up to 50 MB and there is no cap on the page count.
Click Extract text. Every page is read in turn and a counter shows which page it is on, so a long document does not look like it has hung. The text is joined together in reading order and shown to you before anything is saved.
Check the preview before you download
The panel shows the character count, the page count, and the start of the text itself. That preview is the fastest way to tell a real text PDF from a scan: if the words are there, the extraction worked, and only then do you click Download .txt.
Turn the page dividers on or off
By default a line reading Page 1, Page 2 and so on is inserted where each page begins, which makes it easy to trace a quote back to its page. Untick the box for one continuous block instead. With the dividers off, pages that held no text are dropped completely rather than leaving a mystery gap.
Why did I get nothing back
A scanned PDF is a photograph of a document. There are no characters inside it, only pixels that look like characters, so there is nothing to extract and the tool says so instead of handing you an empty file.
The fix is text recognition. Run the scan through OCR to add a real text layer, then bring the result back here.
See also: Make a scanned PDF searchable
Only some of my pages had text
That is normal in a document that mixes typed pages with scanned inserts, such as a contract with a signed page photographed at the end. The tool counts the empty pages and tells you how many there were, so you know exactly which part is missing rather than assuming the whole file failed.
Does it keep columns, tables and layout
It does not. You get the reading order as plain text. Runs of spaces and tabs are collapsed to one space, so a table loses its alignment and a two-column page can interleave the columns, because the words come out in the order the PDF stored them rather than the order your eye reads them.
That is the right trade for feeding text into a search, a script, or a language model. When the layout is the point, convert to a document format instead.
See also: Convert a PDF to Word
Get a PDF table into a spreadsheet
Plain text is the wrong shape for numbers you want to sum. Rows come out as sentences and the column boundaries are gone. Use the spreadsheet converter, which keeps the rows and columns as cells you can sort and total.
See also: Convert a PDF to Excel
Extract text from a password-protected PDF
A locked file cannot be read until the password comes off, so the upload card turns it away and points you at the unlock tool. Remove the password, extract from the plain copy, then put the password back on the original if you still need it.
See also: Unlock a PDF
Copy text out of a PDF that will not let you select it
Some PDFs carry a flag that tells a reader app to block selecting and copying. It is a request, not a lock, and it is not the same as a password. If the text is genuinely inside the file, it comes out here.
Extract text from a long PDF
There is no page limit. A 500 page report works the same way as a two page letter, it just takes longer, and the page counter keeps you posted. The size ceiling is 50 MB per file, and a very long document can push a phone's memory, so a laptop is the better place for the big ones.
Keep accents and other alphabets intact
The .txt is written in a universal encoding, so accented Latin, Cyrillic, Greek, Arabic, Chinese and Japanese all survive exactly as they were in the PDF. Nothing is flattened to plain ASCII and nothing is replaced with a question mark.
Feed a PDF into a script or a search index
Turn the page dividers off, download the .txt, and you have one clean block ready to index, diff, or paste into a prompt. Blank lines are tidied and repeated spaces are collapsed, which is usually exactly what a downstream tool wants.
Ask a question instead of reading the whole file
If the reason you wanted the text was to find one clause or one number, extracting the whole document is the long way round. The question tool reads the PDF and answers directly, pointing at the page it came from.
See also: Ask questions about a PDF
Extract text without uploading the PDF
The document is read inside the browser tab and the .txt is built there. Nothing is sent to a server, so a contract, a payslip or a medical record never leaves your machine, and there is no retention window to trust. There is also no hourly cap and no signup step. The free plan does count finished files, 10 a day shared across every tool.
Extracting text from a PDF, answered
Does my PDF leave my device?
The text is pulled out in your browser and the .txt file is built there too. Nothing is uploaded, which matters for contracts, statements and anything else confidential.
Why did I get no text back?
Your PDF is almost certainly a scan. A scanned page is a photograph of a document, so there is no text inside the file to extract. Run it through the OCR tool first to add a text layer, then convert it here.
Does it keep the original layout?
No, and that is deliberate. You get the reading order as plain text, with runs of spaces collapsed, so columns and table cells lose their alignment. If you need the layout preserved, convert to Word instead.
What are the page dividers for?
They mark where each page began, so you can trace a line of text back to its page in the original. Turn them off if you want one continuous block of text, for example when feeding the output into another tool.
Can I extract text from a password-protected PDF?
Not while it is locked. Remove the password with the unlock tool first, then extract the text.
Is there a page limit?
The free plan covers 10 finished files a day, shared across every tool, resetting at midnight UTC. Files can be up to 50 MB. Long PDFs take longer because every page is read in turn, and the tool shows which page it is on as it works.
Are accents and other alphabets kept?
Whatever the PDF holds comes out intact. The .txt is written in a universal encoding, so accented Latin, Cyrillic, Greek, Arabic, Chinese and Japanese text survives the trip as long as it was real text in the PDF.
Is anything stamped on the output?
No. There is no watermark on any plan, and a .txt file is your text and nothing else. There is no signup either.
Need the words changed, not copied?
Open the PDF in the editor, click any line of text, and type over it.
Edit a PDF nowPDF to JPG
Turn every PDF page into a JPG. Quality selector, ZIP for multi-page.
OpenPDF to PNG
Turn every PDF page into a crisp, lossless PNG. ZIP for multi-page.
OpenJPG to PDF
Combine JPG and PNG images into one PDF. Reorder pages, then download.
OpenWord to PDF
Turn a .docx into a PDF (text, images, tables, links) without leaving your browser.
Open