Find and truly remove personal data from a PDF
Detect emails, SSNs, card numbers and more inside a PDF, redact them so they can't be recovered by copying text, then verify the file is clean.
Your progress
Find and permanently remove personal data from a PDF — real redaction, not black highlighter drawn over live text. 1FileTool's Detect PII scans for emails, phone numbers, SSNs, card numbers, IBANs and IP addresses; Auto-Redact PII (Pro) burns them out of the file; Verify Redaction proves nothing extractable is left.
What you'll need
- 1FileTool installed
- The sample contract below — every identifier in it is intentionally fake (
000-00-0000,[email protected], a Visa test card number) - Pro for the one-click redact step; the Detect and Verify steps are free
redact-pii-from-a-pdf.zip
Sample files · 2.0 KB
- contract-fake-pii.pdf1.9 KB
Protect your originals first
Replace source is ON by default: the output overwrites the original file and no backup is kept. Before following along, open Settings › General and turn Replace source off, or pick a separate output folder.
Steps
1. Run Detect PII
Open Privacy Tools › Detect PII, drop contract-fake-pii.pdf, and press Run on 1 file. The report appears immediately: "6 items found across 1 page" — each match on its own row with a type badge ("Email address", "Social Security number", "Credit card number"…), a masked preview like ja********om, and its page number.
2. Read the report before redacting
Check what was caught before removing anything. If the report says "This PDF has no extractable text, so it is probably scanned", the file is already pixels — jump to the note in If it didn't work.
3. Run Auto-Redact PII
Open Privacy Tools › Auto-Redact PII, drop the same PDF, and press the run button. It re-detects every match and permanently burns each one out: affected pages are re-rendered as images with the matches painted over, so copy-paste can't recover the text. The result is written as redacted_contract-fake-pii.pdf in your output folder.
On the free plan this step stops at a friendly error — "One-click PII redaction is a Pro feature… Scanning for PII is free." If you don't have Pro, the Detect report still tells you exactly what to remove by editing the source document and re-exporting the PDF.
4. Verify the output
Open Privacy Tools › Verify Redaction and drop redacted_contract-fake-pii.pdf. Because redacted pages were re-rendered as images, the answer comes back "Clean — no extractable text remains in this PDF." That is the check to save a screenshot of when compliance asks.
5. Try copying the data out
For the convincing test, open the redacted PDF in any viewer and try to select or search the email — there is nothing selectable where it used to be. Drawn-on boxes in ordinary editors leave the text layer intact underneath; this does not.
The result
The sample contract goes in with six personal identifiers and comes out as redacted_contract-fake-pii.pdf with zero extractable matches — verified by the app's own Verify step, not by eyeballing the black boxes.
If it didn't work
- "This PDF has no extractable text" — it's a scan or photo. PII detection reads the text layer; for image-only PDFs, redact regions visually in your PDF editor, or convert pages with OCR first.
- "No personal data found" — good news, or a format the detector doesn't cover (it knows emails, phones, SSNs, card numbers, IBANs and IPs — not names or addresses).
- "PRO_REQUIRED" / the upsell error: Auto-Redact needs Pro. Free users still get the full Detect report and can strip everything by re-creating the PDF.
- A match survived: re-run Detect on the output — Verify says "N items are still extractable" if anything remains, so you'd know.
Try next
Remove GPS and camera data covers the image side of the same problem; Privacy Tools › Strip Metadata also scrubs author/title metadata off PDFs before you share them.