How to Extract Text from Images and Scanned Documents (OCR)

Getting editable text out of a photo or scan is easier than it used to be β€” here's how to do it with tools you likely already have.

  1. Use your smartphone camera's built-in text recognition

    Most modern phone camera and photo apps can detect text directly in a live camera view or a saved photo, letting you select, copy, and paste it without a separate app.

  2. Try Google Lens for a quick one-off extraction

    Point Google Lens at text (in the camera or on an existing image) and it recognizes and lets you copy the text instantly, translate it, or search it β€” useful for a fast single extraction without setting anything up.

  3. Use a free browser-based OCR site for documents or PDFs

    For scanned documents or PDFs rather than a single photo, a dedicated OCR website can process a whole file at once and export the result as editable text or a searchable PDF.

  4. Use OCR features built into office software

    Programs like Microsoft OneNote and Google Docs (via "Open with Google Docs" on an image or scanned PDF) include built-in OCR that can extract text as part of a workflow you're likely already using.

  5. Improve the image quality before running OCR if accuracy is poor

    OCR accuracy drops sharply with blurry, tilted, low-resolution, or low-contrast images. Retaking the photo in good lighting, holding the camera straight, and avoiding glare noticeably improves results.

  6. Proofread the extracted text

    OCR occasionally misreads similar-looking characters (like "0" and "O," or "l" and "1"), so always review the output, especially for numbers, names, and anything you plan to use verbatim.

How OCR actually works

OCR software does not "read" text the way a person does β€” it analyzes the shapes and patterns in an image, compares them against known character shapes, and predicts the most likely letter or number for each shape. This is why OCR performs very well on clean, printed text in a common font, and considerably worse on handwriting, stylized fonts, or low-quality scans, where the shape patterns are less predictable.

Choosing the right OCR tool for the job

A quick camera-based tool like Google Lens is ideal for grabbing a short piece of text on the fly β€” a sign, a quote, a phone number. For processing many pages or an entire scanned document at once, a dedicated OCR tool or office-software feature designed for bulk documents will generally be faster and more accurate than repeating a single-photo tool page by page.

Frequently Asked Questions

Why does OCR sometimes get certain characters wrong even in a clear photo?

Certain characters look nearly identical in many fonts β€” like a capital "I," lowercase "l," and the number "1" β€” so OCR occasionally confuses them even in an otherwise sharp image. Reviewing and correcting the output is standard practice, not a sign the tool failed.

Can OCR extract text from handwriting?

Basic OCR is built for printed text and struggles significantly with handwriting. Some newer tools include dedicated handwriting recognition, but accuracy is still noticeably lower than for printed text, especially with less tidy handwriting.