Make a Scanned PDF Searchable (Beta)

Add a hidden text layer to a scanned PDF so you can search it, select text and copy it, with the pages kept exactly as they are.

Opens: PDF, JPG, PNG, TIFF, WebP, BMP, HEIC, AVIF
Saves: PDF, TXT

Read the text

Scanned PDFs or pictures
Choose files… or drop them here
.pdf, .jpg, .jpeg, .jpe, .jfif, .png, .tif, .tiff, .webp, .bmp, .heic, .heif, .hif, .avif · up to 200 MB each, 200 files max
    Add more files

    PDFs are read page by page and each gives a searchable PDF. Pictures (photos, screenshots, scans; several or a multi-page TIFF) become one PDF, in the order you add them.

    Options

    Choose the language the text is written in: letters with accents and whole words are read much better.

    For documents that mix two languages, such as English terms in a Russian letter. Reading takes about a quarter longer.

    Each page's text direction is checked, and pages scanned the wrong way round are turned. The page pictures themselves are not changed.

    Pages scanned at a slight angle are straightened. Their pictures are then saved again, so the PDF can get larger.

    Typed PDFs and pages read before have text you can select already.

    More options

    Leave empty for all pages, or write pages like 1-3, 5, 8- (8 to the end). The other pages stay in the PDF as they are.

    Only for pictures: how they become PDFs.

    Your files are deleted after processing.

    A scanned PDF opens like any other, but press Ctrl+F and nothing is found: every page is a picture. Making it searchable means reading the words on each page and adding them as text the PDF reader can find, without changing how the pages look.

    How to make a scanned PDF searchable

    1. Add the PDF above (or several).
    2. Choose the language of the text, and a second one if the document mixes two.
    3. Press Read the text and download the searchable PDF.

    Open it in any PDF reader: Ctrl+F finds words, you can select and copy sentences, and search tools on your computer or in document management systems can index it.

    What changes and what doesn't

    The text goes into an invisible layer placed exactly over the words of each page. The page pictures are kept as they were scanned: no recompression, no cleaning, no change in quality, so the file stays about the same size. Pages scanned sideways or upside down are turned upright by setting the page's rotation, not by redrawing it.

    Pages that already have text, such as typed pages in a PDF that also has scans, are left as they are and only the scanned pages are read. If a PDF was read by another OCR program and its text is poor, choose Replace earlier OCR text: the old invisible layer is removed and the pages are read again, while typed text stays.

    Pitfalls

    • The language matters. A German letter read as English loses its umlauts and many words.
    • Crooked scans are read, but straightening them (an option) gives better results on pages that are more than a degree or two off. It saves the page pictures again, so the file can grow.
    • Password-protected PDFs can't be changed here; remove the password in a PDF program first.
    • Very large pages, such as maps or posters scanned at high resolution, are kept but left without text above 100 megapixels.

    More about what this tool reads and writes: OCR PDF.