PDF Metadata: What's Hidden in Your Files

Every PDF you've ever shared carried a stowaway: metadata. Author names, the software that made it, creation dates, sometimes the editing history and even deleted content. Most of the time it's harmless. Sometimes — a job application revealing a previous employer's name, a 'clean' contract showing its true author — it's a quiet disaster.

Here's what's actually inside your PDFs, how to inspect it, and how to strip it before sharing.

Try it free: Compress PDF

Strip metadata from your PDFs — free, in your browser.

Open Compress PDF

What metadata lives in a PDF

PDFs carry two metadata systems. The classic Document Information Dictionary holds Title, Author, Subject, Keywords, Creator (the app that made it), Producer (the PDF engine) and creation/modification dates. The newer XMP stream can hold far more: edit histories, embedded thumbnails, camera data from scanned images, and custom fields added by authoring software.

Beyond declared metadata, PDFs can leak through structure: incremental saves preserve earlier versions of edited content, and un-flattened annotations or form fields may contain data you thought you removed.

Real ways metadata burns people

  • A resume's Author field naming a current employer the applicant didn't mention
  • A legal document's edit history revealing negotiation positions
  • Scanned images carrying GPS coordinates from the phone that photographed them
  • 'Final' reports containing earlier draft text recoverable from incremental saves
  • Creator fields advertising outdated, vulnerable software versions

How to inspect your PDF's metadata

The quick check: open the PDF in any reader and look at Document Properties (Ctrl+D / Cmd+D in most readers) — Title, Author, and app info are right there. For the full picture including XMP, the free command-line tool ExifTool dumps everything: exiftool yourfile.pdf reveals fields readers never show.

Make this a habit before sending anything sensitive externally. It takes ten seconds.

How to strip it

Re-saving a PDF through a clean serializer clears the standard Document Information fields — this is exactly what free browser compressors do as a side effect of rewriting the file. For thorough sanitization including XMP, incremental-save history and hidden layers, use Acrobat's 'Sanitize Document' feature or ExifTool's metadata removal.

One caution: stripping is destructive to the metadata only, never to visible content. But verify afterward — open properties and confirm the fields are actually gone.

When to keep metadata

Metadata isn't always the enemy. Archival workflows, digital asset management and accessibility all rely on good Title/Author/Language metadata. The rule: keep rich metadata for files you manage, strip it for files you share externally — especially with strangers.

Scanned vs born-digital: different leaks

Born-digital PDFs (exported from Word, InDesign) leak authoring metadata: usernames, software versions, edit histories, tracked changes. Scanned PDFs leak capture metadata instead: the scanner or phone model, GPS coordinates from camera photos, and timestamps accurate to the second.

The defenses differ accordingly. For born-digital files, sanitize authoring fields and flatten tracked changes before export. For scans, strip EXIF from the images before assembling the PDF — or run the finished PDF through a metadata scrubber. Check both kinds with ExifTool before sharing.

Related guides

Frequently asked questions

What personal data can leak through PDF metadata?

Author names, organization, software used, creation and edit dates, GPS coordinates in scanned images, and sometimes earlier draft content preserved by incremental saves.

How do I view a PDF's metadata?

Open Document Properties in any PDF reader for the basics (Title, Author, app info). For the complete picture including hidden XMP data, use the free ExifTool.

How do I remove metadata from a PDF for free?

Re-saving through a clean serializer — which free browser PDF compressors do automatically — strips standard metadata fields. For deep sanitization use Acrobat's Sanitize Document or ExifTool.

Does compressing a PDF remove its metadata?

When compression rewrites the file cleanly, yes for standard fields (author, title, producer). It may not clear XMP streams or incremental-save history — use a dedicated sanitizer for sensitive documents.

Private by design: SwiftPDF tools run entirely in your browser — your files never leave your device.