Remove metadata from a PDF

Every PDF records things you never typed: an author name taken from the account that made it, the software that produced it, when it was created and last changed, and often an embedded metadata packet holding the same details again. This guide lists all of it and shows how to strip what you choose.

Remove metadata from PDF

What a PDF says about you

The document properties hold a title, an author, a subject and keywords, plus the name of the program that created the file and the name of the program that produced it. The author is usually your account name, and the producer usually names your software and its version.

Alongside those sit a creation time and a modification time, both accurate to the second. On a document you claim to have written last week, a creation timestamp from this morning is the detail that gives it away.

The parts people miss

Many PDFs also carry an XMP metadata packet, which is a separate block holding much the same information in another format. Clearing the document properties in an ordinary reader often leaves that packet behind, so the data you thought you removed is still readable.

There is also a document identifier stored in the file, and there may be attachments, which carry their own file names. A spreadsheet attached to a report tells anybody who looks what it was called and where it came from.

See it before you remove it

Open the privacy scan and choose the file. It lists what is actually there, grouped as document details, software, dates, embedded files and other metadata, with the value of each field shown so you can read it.

Seeing the values is the point. People are surprised by which fields carry their real name, an old employer, a template author or a file path, and you cannot make a sensible decision about a field whose contents you have not seen. The metadata removal route covers this job and points at the same tool until its own preset ships.

Remove what you choose

Select the entries to strip and save. You choose per item rather than all or nothing, which matters when a title is genuinely useful and an author name is not. At least one item has to be selected, since there is otherwise nothing to do.

The tool then does two checks. It compares every page of the output against the input, so the document itself is unchanged, and it reopens the file and confirms that each item you selected is genuinely gone rather than merely blanked.

The side effect worth knowing

Merging documents and organizing pages both clear the document properties as part of their work, because each builds a fresh document. A file that has been through either of those carries no title, author, producer or dates from its source.

That is not a substitute for the privacy scan, which is deliberate and shows you what it found. It does mean a merged file is often already clean of the fields people worry about, and it is worth checking rather than assuming either way.

What this does not reach

The scan reads what the PDF records about itself. It does not open the pictures inside the pages, so camera and location data living inside an embedded photograph is a separate layer. Clear that in your photo library before the images become PDF pages.

It also does not touch content that is visible on the page. A name printed in a footer, a signature block or a header is page content, and removing that permanently is redaction, which destroys the content rather than hiding it.

Before you publish

Run the scan on the file you are actually sending, not on the version you edited three steps ago. Every save through every program rewrites the producer field and the modification time, so the last export is the one that matters.

Encrypted files are refused, so unlock a copy first if the document is protected. The blank page guide covers tidying a document before it goes out, the privacy page explains what this site records, and the pages hub lists the rest of the cluster.

Questions people ask

What personal data does a PDF actually store?
A title, author, subject and keywords, the creating and producing software, creation and modification timestamps, a document identifier, and often an XMP packet repeating those details. Attachments carry their own file names as well.
Why is my name still in the file after I cleared the properties?
Usually because the XMP metadata packet was left behind. It holds the same details in a separate block, and clearing the properties dialog in a reader often does not touch it.
Can I keep the title and remove only the author?
Yes. You select the entries to strip one by one, so a useful title can stay while the author name goes. At least one item has to be selected for there to be anything to do.
Does stripping metadata change the pages?
No. The output is compared page by page against the input, and the tool separately confirms that each selected item is gone rather than blanked. The document itself is untouched.

Do it now

The tool runs in this browser. Your file never leaves the machine, and the result is checked before you download it.

Remove metadata from PDF

Where to go next