OCR a scan to text (English)

Extract text from a scanned PDF with offline OCR. English recognition only, output is a plain .txt file — here's exactly what to expect.

The OCR tool in 1FileTool turns a scanned PDF into a plain text file on your Mac — no upload, no cloud OCR service. It recognises English text only and writes a .txt file, not a searchable PDF.

Protect your originals first

Replace source is ON by default: the output overwrites the original file and no backup is kept. Before following along, open Settings › General and turn Replace source off, or pick a separate output folder.

Run OCR on a scan

  1. Open PDF Tools › OCR (tool page).
  2. Drop the scanned PDF.
  3. Run it. Each page is rendered at 300 DPI and recognised, so a long scan takes a while — progress is per page.
  4. Find <name>_ocr.txt in your output location. Pages are separated by a --- Page Break --- line.

What you get (and what you don't)

  • The output is a plain .txt — words and line breaks, no layout, no fonts. A two-column scan reads straight through.
  • Recognition is English-only. Pages in other languages come out garbled or empty.
  • Clean, flat scans at reasonable contrast recognise well. Handwriting, stamps, skewed photos and faint thermal-paper text will need manual cleanup.
  • OCR does not produce a searchable PDF — it extracts text. To get editable paragraphs in Word instead, use Scanned PDF to Word.

If you need the text inside another format

  • Scanned PDF to Word runs the same OCR but writes a .docx of plain paragraphs — see PDF to Word.
  • PDF Tools › Extract Text copies the embedded text layer of a digital PDF — instant, but it finds nothing on a scan because a scan has no text layer.

Tools used

Keep the next file job off the internet.

Every tool is in the free download. Upgrade once, when the daily limit starts getting in your way.

PrivateOfflinePay once

No account, no credit card. Pro includes a 7-day money-back guarantee.