Why PDFyre exists
PDFyre started with a very personal problem: an 890-page medical textbook that had been scanned into a PDF and could not be searched. Every word I needed to find was trapped inside images, and no free online tool would touch a file that big β let alone keep it private.
The story
I'm a medical student. My life is PDFs β lecture slides, textbooks, research papers, hospital forms. During exam season I got a scanned copy of a core textbook and hit a wall: I could not search it, copy from it, or jump to any topic. Ctrl+F found nothing, because there was no text in the file, only pictures of pages.
Online OCR tools had two problems. First, they capped file sizes, and my book was 132 MB β far over every limit I found. Second, they wanted me to upload a document full of personal notes and patient-related study material to a server I did not control. That was a non-starter.
So I built the tool I wished existed: OCR that runs entirely in the browser. The PDF is read, processed, and re-saved on your own device. Nothing is uploaded, ever β there is no server waiting to receive your files. The 890-page book that started all of this was the first file it processed successfully.
What PDFyre stands for
Privacy first
Every tool runs locally on your device. Your documents never leave your hands.
No limits
No file size caps, no queues, no account, no watermarks. If your device can handle it, PDFyre can.
Works anywhere
Open it in any modern browser on any device β no software to install, nothing to pay.
Always free
Every tool is free. The site is supported by advertising, not by selling access to your files.
The tech behind it
PDFyre uses open, battle-tested technology β PDF.js for reading and rendering PDFs, pdf-lib for building them, and Tesseract.js for OCR β all executed in your browser's own engine.
Built for people who just want it to work
Try it with your own files β searchable, compressed, merged, split, all in your browser.
π₯ Open PDFyre