Extract Text from PDF
Copy the text out of a PDF, or save it as a plain text file. Works with any language that the PDF stores as real text, including Hindi.
Click to upload a PDF, or drag and drop
PDF files only, up to 50 MB
Options
Extracted text
Your PDF is processed entirely in your browser and is never uploaded anywhere.
Frequently asked questions
How do I copy text from a PDF?
Upload the PDF, choose the pages if you do not need all of them, and click Extract text. The text appears in a box where you can edit it, copy it to the clipboard or download it as a .txt file.
Why does it say no text was found?
Some PDFs are made of pictures, for example pages that were scanned or photographed. They look like text but contain none. To read those you need OCR (optical character recognition). The OCR Wizard in Nyay Sahayak lets you select a region of a scanned page and turn it into editable text.
Will Hindi and other languages work?
Yes, if the PDF stores the words as text. The tool reads whatever text the file contains. Occasionally an older PDF uses a custom font that maps letters to the wrong characters, and the extracted text then looks garbled. That comes from how the PDF was made, and OCR is the usual workaround.
What does “Join lines into paragraphs” do?
A PDF stores each line of a paragraph separately. This option joins lines that belong together, and re-attaches words that were split with a hyphen at the end of a line, so the text flows when you paste it into a document.
Is my PDF uploaded to a server?
No. The text is read in your browser, so the document never leaves your device. Password-protected PDFs cannot be opened.