Every PDF tool you need,
running inside your browser.
Compress, convert, merge, sign, redact and read scans. Choose a tool and open a file — it is processed on this device and never uploaded.
Why this is different
No document is uploaded. This can be independently verified.
Comparable services transmit your file to a server and undertake to delete it subsequently. PaperJET has no server to which a document could be sent: the tools are compiled to WebAssembly and execute within this browser tab, on your own device. Nothing is queued, nothing is capped, and the work continues if you go offline after the page has loaded.
Questions
Frequently asked questions
Are documents uploaded to a server?
No. Every tool executes within your own browser using WebAssembly. The file is read from local storage by the browser and is not transmitted, so no server holds a copy of it. This may be verified by disconnecting from the network once the page has loaded; the tools continue to function normally.
Is PaperJET free to use?
Yes. All tools are provided free of charge, with no account required, no watermark applied and no file size limit.
Can a scanned PDF be processed?
Yes. The OCR tool reads the text from a scan and can return the same scan with an invisible text layer applied, so that the document appears unchanged while becoming searchable and available for copying.
Can a PDF be converted to Word or Excel?
Yes, where the PDF already contains text. Where the file is a scan, the OCR tool should be used instead; the interface will indicate which case applies.
Is there a file size limit?
No. As processing is performed on your own device rather than on a server, the constraint is the available memory of that device rather than an upload cap.
Read the text from a scanned Hindi document
A scanned Hindi document is a picture of a page: the words are visible but there is no text in the file to search, copy or select. This reads the Devanagari on the page and puts the text back, using a Hindi language model chosen for you. Recognition runs on this device, which is worth knowing when the scan is a certificate, a land record or anything else you would rather not hand to a website.
How to read the text from a scanned Hindi document
- Open the scanned PDF. Hindi is already selected as the language.
- Run the recognition. The Hindi model is around 18 MB and downloads once on first use, then stays cached.
- Save the result. The scan looks exactly as it did, with the recognised text applied as an invisible layer beneath it.
The recognised text is placed behind the original image rather than replacing it. Where recognition is imperfect — and on Devanagari it sometimes is — the page still looks exactly as it was scanned, and nothing has been lost.
Devanagari is harder to recognise than Latin script, because the headline joining the letters and the matras above and below them carry meaning that blurs easily. A straight, evenly lit scan at 300 DPI reads far better than a phone photograph at an angle.
Both are written in Devanagari, but the language model expects different words. For a Marathi document, use the Marathi option instead — the script will look right either way, but the accuracy will not.
Questions about Hindi OCR
Is the Hindi document uploaded for recognition?
No. The recognition engine and the Hindi language model both run in this browser tab. The document is read from your own device and never sent anywhere.
Why is the first run slow?
The Hindi language model is roughly 18 MB and is fetched once, on first use. It is cached afterwards, so subsequent documents start immediately.
How accurate is Hindi recognition?
Good on a clean, straight scan of printed Hindi; noticeably worse on faint print, handwriting or a photograph taken at an angle. Devanagari is more demanding than Latin script because the matras above and below the line matter.
Can it handle a document with both Hindi and English?
Yes — choose the combined English and Hindi option in the language list on the OCR tool, which loads both models together.