A box of old paperwork is only useful if you can find something in it. Digitizing it solves storage, but digitizing it and making it searchable is what actually lets you find “that one receipt from March” or “the certificate from 2019” without flipping through a folder by hand.
Digitizing paper is really two separate jobs, and it’s worth keeping them separate: capture turns the physical page into a digital file, and making it searchable turns the text in that file into something you can search, select, and copy, not just look at. Skipping the second step is the most common reason old scans end up just as hard to search as the paper they replaced — you’ve digitized the picture of the page, not the words on it.
For capture, go to the Scan to PDF page and photograph each page with your phone’s camera — there’s no dedicated scanner required, the tool works from ordinary photos, cropping and cleaning up each page as you go, so a stack of paper becomes a single PDF. Between capture and recognition is the point to clean up: a batch captured quickly will have a page or two upside down, a duplicate shot, or a blank sheet that got photographed by accident, and fixing that before running OCR matters, since recognition is only as good as the page it’s given and a sideways page recognizes badly or not at all. Edit PDF rotates, deletes, and reorders pages from one thumbnail view, so a whole batch is corrected in a single pass and a single download. Then, for the searchable part, go to the OCR PDF page, upload the file, and let it recognize the text on every page — this adds an invisible text layer under the image without changing how the page looks, so you end up with a document that looks exactly like the original but responds to Ctrl+F.
This approach is worth it most for old personal records — certificates, IDs, contracts — that you want backed up and findable without digging through a filing box; handwritten or typed notes you want to search by keyword later rather than remembering which notebook a specific note is in; receipts and paperwork for taxes or expense reports (see How to Scan Receipts for Expense Reports for a workflow built specifically around that case); and anything currently only backed up as “a photo in my camera roll,” which is searchable by date but not by content.
One caveat worth knowing up front: OCR is reliable on typed and printed text, but handwriting is a much harder case — recognition quality depends heavily on how neat and consistent the writing is, and results are inconsistent even on clear handwriting. For handwritten pages, treat OCR as a bonus rather than something to depend on, and keep the original scan as the reliable copy either way.
Once you’re through a batch: combine multiple scans into one PDF if you captured pages as separate images before starting, compress the result if you’re digitizing a large batch and file size adds up, or password-protect it if the documents contain anything sensitive.
Two steps get one box of paper digitized. Doing it continuously, so paper never piles up again, is a different problem — one covered in How to Create a Paperless Home Office for household paperwork and How to Archive Business Documents Digitally for company records.
Start with the free Scan to PDF tool, then run OCR PDF — both free, no sign-up, and files wiped the moment each download finishes.