Your privacy choices

Paperuna uses necessary browser storage to remember your choices. With your permission, we may also use analytics and advertising technologies. Optional technologies stay off until you choose.

Privacy and security

How to inspect and remove metadata from a PDF

Learn which identifying properties PDFs may contain, how to remove common metadata, and why metadata cleaning is not a complete forensic guarantee.

By Aaron McGuireUpdated 6 minute read

What PDF metadata is

Metadata is information describing a document rather than its visible page contents. Common PDF properties include title, author, subject, keywords, creator application, PDF producer and creation or modification dates. A document exported from an office application may therefore identify a person or organisation even when those details do not appear on the page.

PDF readers expose some properties through a document-information panel, while other values may sit in XML metadata or embedded objects. The exact contents depend on the software and workflow that created the file.

When metadata matters

Metadata can be useful for organising archives and identifying authoritative versions. It can also disclose more than intended when a draft, anonymous submission, public-information response or client document is shared externally.

Removing metadata should be a deliberate decision rather than an automatic belief that all metadata is dangerous. Dates and provenance may be important for records management, evidence or accessibility workflows. Always keep an original when those properties may need to be retained.

Common fields to remove

A basic cleaning process can clear the author, title, subject, keywords, creator, producer and modification properties, then save a separate output. Rename the downloaded file if its filename itself contains private information.

  • Author or organisation name
  • Internal document title or subject
  • Keywords and classification labels
  • Creator and producer software
  • Creation and modification timestamps

Why cleaning metadata has limits

Clearing common properties is not the same as forensic sanitisation. PDFs may contain comments, form values, embedded files, scripts, hidden layers, image metadata, previous content left by a particular editing workflow or information visible only when objects are inspected.

Paperuna’s metadata tool removes common document properties where supported. It does not claim to discover every possible hidden data element in every PDF. For highly sensitive or regulated disclosure, use an appropriate specialist review process.

A practical pre-sharing checklist

Open the cleaned result as a separate file and inspect its document properties. Review comments, attachments and form fields, search for confidential terms, and check the visible pages. Confirm that the filename is suitable and share the cleaned copy—not the working original.

If visible content must also be removed, use permanent redaction before the final review. Metadata removal cannot make a name printed on a page disappear.

Common questions

Does changing the filename remove PDF metadata?

No. The filename and the metadata stored inside the PDF are separate.

Does metadata removal redact visible text?

No. Use a permanent redaction workflow for sensitive content that appears on a page.

Try it privately

Put this guide into practice

Paperuna processes supported documents locally in your browser. Your file is not uploaded for processing.

Remove PDF metadata

This guide provides general information, not legal, regulatory or professional advice. Review important outputs and keep an untouched copy of the original document.