How to redact a PDF

Redacting a PDF means deleting the marked content, not drawing over it. Mark every private region, rebuild each affected page, keep only the unmarked text you still need, then inspect the written file for live text and hidden data. The pdfpix tool does those steps in this browser tab and refuses a result that fails its checks.

Redact PDF

How to redact a PDF permanently

  1. Choose the final PDF

    Finish merging, compressing and signing first, then choose the exact file you plan to send.

  2. Mark every private region

    Draw boxes, select text blocks, find repeated text, or review the in-tab detector suggestions page by page.

  3. Choose the text mode

    Keep unmarked text selectable, or flatten every marked page when removing all selectable text there is acceptable.

  4. Read the result checks

    Confirm the black regions, text layer, removed raw fragments and private metadata all pass before downloading.

A cover is not a redaction

PDF pages are lists of drawing instructions. Adding a black rectangle puts one more instruction above the name, number or image you meant to remove. The covered material can remain in the content stream, text layer or an annotation below it. A person may still select it, search for it or extract it without removing the rectangle.

A real redaction makes the source content absent from the result. Redact PDF rebuilds each marked page from a fresh render and applies the black region to that new page. It does not copy the original marked page content into the output. That is the difference between using a PDF redaction tool and colouring over a page.

Finish the document before redacting it

Work on the final copy. Merging can reorder pages, compression can rebuild a document, and a later signature or stamp can write new objects and metadata. If any of those happens after redaction, the file you checked is not the file you send.

Keep an untouched source in a separate place. The redacted copy is for distribution; the source is for correction if you missed a region. Do not use the redacted file as the only record, because removal is meant to be irreversible in that copy. Rename the finished file too. A revealing file name sits outside the PDF and no redaction engine can change it.

Find page content and data outside the page

Read every page once and mark names, addresses, identifiers, signatures, faces, barcodes and any surrounding words that reveal the same fact. A label can be as sensitive as its value. Removing an account number while leaving “retirement account” beside the mark may still disclose more than you intend.

Then inspect what is not printed. A PDF can carry an author, title, creator, producer, creation date, modification date, XMP packet, attachments, annotation text, form values, outlines and open actions. Page redaction and PDF metadata removal are different jobs. The metadata guide covers the second pass.

Mark selectable text and scanned text differently

Selectable text gives you three useful shortcuts. Select the whole text block, search for a repeated value across pages, or let the in-tab detector suggest people, street addresses and validated identifiers. Suggestions are a review queue, not an all-clear. They can miss private information, so redacting a PDF still ends with a page-by-page read.

A scan is a picture. Searching for text will not find words that exist only as pixels, and a text detector cannot mark them. Draw a box around the printed area instead. Include the full height of letters and a small margin on every side. If the same scan also has an OCR layer, the mark must cover both what you see and the text coordinates below it.

Decide what should remain selectable

The default mode keeps unmarked text selectable. It rebuilds the marked page as a render, then adds back a text layer containing only characters outside the redaction marks. This is the practical choice for reports, filings and long documents that still need search and copy after private lines are removed.

The strict mode flattens every marked page and leaves no selectable text on that page. It is the narrower safety choice when nobody needs to search the surviving text. Unmarked pages remain unchanged in either mode. The trade is explicit because turning a whole document into images would break accessibility and make clean text harder to use.

Check the file rather than the preview

A page preview proves only that black pixels are visible. The output needs several separate checks: each marked region renders black, no selectable text overlaps a mark, removed text fragments are absent from the raw PDF bytes, and supported private metadata is gone. The tool reopens the output and runs those checks before it shows a download.

Open the downloaded file again in a different reader. Search for every exact name and number you removed, try to select through each mark, and inspect the document properties. For a second view, Redaction check looks for supported opaque cover boxes with live text underneath and reports the text it can recover.

Know what the checker cannot promise

No automatic scan sees every possible leak. Text burned into an image has no selectable layer to compare. A cover made from an unusual clipping path or many small shapes may not look like one rectangle to the parser. A hidden fact can also survive in a filename, a previous email attachment or a cloud version history outside the PDF.

That is why the last pass belongs to a person who knows what the document contains. Review every page at normal size and at high zoom. Read the checker's limits. Send only the copy you inspected. The recovery guide explains how fake redactions, crops and hidden layers differ from content that was actually removed.

A short final checklist

If you need to redact information in a PDF today, use the tool first and this list as the review pass. If you need to redact text in a PDF that is only a scan, use drawn boxes rather than search. If the job is described as “black out a PDF,” the same rule holds: the content has to be destroyed, not hidden behind black ink.

Questions people ask

How do I redact text in a PDF permanently?
Mark the text, rebuild the affected page, and check the written output for selectable text and raw fragments below the mark. Drawing a rectangle in an editor changes the appearance but may leave every character in the file.
Can I redact information in a scanned PDF?
Yes, with drawn regions. A scan stores words as image pixels, so text search cannot find them. Draw a box around the full printed area and review the rendered output at high zoom.
Does PDF redaction remove metadata too?
Page content and metadata are separate. The pdfpix redaction pass strips supported metadata and attachments, but you should still inspect the final file with Privacy scan and rename it before sending.
Should unmarked PDF text stay selectable?
Usually. The default mode keeps only unmarked characters in a new text layer. Strict mode removes all selectable text from marked pages when search and accessibility on those pages are less important than that narrower result.
Can an automatic detector find every item to redact?
No. It can suggest supported people, addresses and validated identifiers in selectable text. It can miss private information and cannot read text that exists only inside a scanned image.
How do I check a redacted PDF before sending it?
Reopen the downloaded file, search for the removed values, try to select through every mark, inspect its properties and run Redaction check. Read the limits and review every page yourself.

Do it now

The tool runs in this browser. Your file never leaves the machine, and the result is checked before you download it.

Redact PDF

Where to go next