Guides

What Metadata Is Stored in a PDF?

A PDF stores metadata in two places: a Document Information dictionary (the classic Title, Author, Subject, Keywords, Creator, Producer and the creation and modification dates) and an XMP packet (an XML block that can hold the same fields plus more). Most of this is invisible in a normal reader but travels with the file. It’s how a document you thought was anonymous can still name you, your software, and exactly when you wrote it. Here’s what’s actually in there.

This guide lists the standard metadata fields, explains the two storage locations, and shows how to read them with the PDF metadata extractor in your browser.

The two places metadata lives

  • The Document Information dictionary (the “Info” dictionary). The original PDF metadata store — a simple set of key/value fields. Every PDF reader can show it under “Document Properties”.
  • The XMP packet. A newer, XML-based metadata block (Adobe’s Extensible Metadata Platform) embedded in the file. It can duplicate the Info fields and carry much more — rights, tool history, custom fields. When the two disagree, modern software generally trusts XMP.

A thorough metadata reader checks both, because a field can be blank in one and filled in the other.

The standard fields

  • Title — the document’s title (often not the filename).
  • Author — frequently your real name or account name, set automatically by the software.
  • Subject and Keywords — description and tags, if the author added them.
  • Creator — the application the document was authored in (e.g. “Microsoft Word”).
  • Producer — the library or app that wrote the PDF (e.g. “Acrobat Distiller”, “Skia/PDF”, a specific converter). This often reveals the exact software and version.
  • CreationDate and ModDate — when the PDF was created and last modified, usually down to the second and time zone.

What this can unintentionally reveal

Because most of these fields are set automatically, a PDF can leak more than its author intended:

  • Who made it — the Author field often carries a real name or corporate username.
  • What software and version — the Producer/Creator pair fingerprints the toolchain, which matters for both privacy and document forensics.
  • When it was made and edited — the timestamps can contradict a document’s claimed date, or show it was edited after “final”.

None of this shows on the page, which is exactly why it’s worth checking before sharing a sensitive document.

How to read a PDF’s metadata

Drop the file into the PDF metadata extractor: it reads both the Info dictionary and the XMP packet and lists every field it finds, in your browser, with the file never uploaded. That last point matters — you’re inspecting a document precisely because it might be sensitive, so it shouldn’t be sent to a server to do so.

Once you’ve seen what’s there, the next question is usually removing it before you share — covered in how to remove hidden metadata before sharing a PDF. For metadata in other file types, see EXIF in photos and the Office metadata extractor.

Frequently asked questions

What metadata does a PDF contain?
Title, Author, Subject, Keywords, the Creator and Producer software, and creation and modification dates — stored in the Info dictionary and/or an XMP packet inside the file.

Where is metadata stored in a PDF?
In two places: the Document Information dictionary (the classic fields) and an XMP metadata packet (XML). A field can appear in one, the other, or both.

Can a PDF reveal who wrote it?
Often, yes. The Author field is usually filled automatically with a real name or username, and the Producer field reveals the software used.

How do I see a PDF’s hidden metadata?
Open it in a PDF metadata extractor, which reads both the Info dictionary and the XMP packet and lists every field.

Is PDF metadata visible on the page?
No. It’s stored in the file’s structure, not drawn on the page, so it stays with the document invisibly unless you inspect or remove it.

Last updated: 16 August 2026.

About Abrar

Abrar builds EasyExtract's free, browser-based extraction tools and writes these guides on getting data out of files — PDFs, spreadsheets, images, archives and Office documents. Every tool runs entirely in your browser, so nothing you open is ever uploaded.

Keep reading