PDF metadata is information about a document stored inside the file but not shown on its pages: the title, the author's name, the software that created it, creation and modification dates, and sometimes more. It helps search engines, document management systems, and viewers organize files. It can also quietly reveal details you never meant to share, such as the name of the colleague who drafted a proposal or the internal title of a project. Removing or editing metadata before you send a file outside your organization is a simple habit that closes that gap.
This guide covers what metadata a PDF can hold, how to see it, when it matters, and how to clean it.
Where metadata lives in a PDF
A PDF has two main places for document-level metadata.
The document information dictionary
This is the classic set of properties that has existed since early versions of PDF. It typically includes:
- Title, which some viewers show in the window or tab instead of the file name
- Author, often filled in automatically from the user account of the creating software
- Subject and Keywords
- Creator, the application that made the original document, such as a word processor
- Producer, the software that converted or wrote the PDF
- Creation date and modification date
XMP metadata
XMP (Extensible Metadata Platform) is an XML-based format embedded in the file. It can repeat the same information as the document information dictionary and add much more, including editing history entries, document identifiers, rights information, and custom fields added by specific applications. Photo and design software in particular can write detailed XMP.
Other places information hides
Metadata in the strict sense is the above, but a PDF can carry other revealing information:
- Images inside the PDF may contain their own embedded metadata, such as camera model or location data in photos, depending on how they were inserted.
- Comments and annotations, which carry author names and timestamps.
- Bookmarks with internal section names.
- Form fields with default or previous values.
- Attachments embedded in the file.
- Application-specific private data, which some creators embed to allow round-trip editing.
Keep these in mind when you think about what a file could reveal.
How to view a PDF's metadata
You can check what a file contains with tools you already have.
- In most PDF viewers, open File, then Properties or Document Properties. The Description tab usually lists title, author, subject, keywords, creator, producer, and dates.
- In web browsers, the built-in PDF viewer often has a document properties option in its menu.
- On macOS, Preview shows basic information through Tools, then Show Inspector.
- On Windows, right-clicking the file in File Explorer and choosing Properties, then Details, can show some fields, depending on installed handlers.
Take a look at a few PDFs you have sent recently. It is common to find an author name you did not expect, a title left over from a template, or a producer field naming a specific internal tool.
Why metadata matters when sharing
Most of the time metadata is harmless. It becomes a problem in a few recurring situations.
Anonymity and attribution
If a document is supposed to be anonymous, such as a peer review, a whistleblower report, a blind job application sample, or feedback to an employer, the author field may name you directly. Even a user account name like a first initial and surname can identify someone.
Internal details
Titles, subjects, and keywords are often filled in while a document is being drafted and never updated. A proposal sent to one client might still carry a title referencing another. File names of earlier drafts or project codes can also appear.
Timeline information
Creation and modification dates can show when a document was really prepared. That may be irrelevant, or it may be sensitive, for example in negotiations or when a document is presented as recent.
Software fingerprints
Creator and producer fields reveal which applications and sometimes which versions an organization uses. That is rarely a concern for individuals, but some security teams prefer not to advertise internal software details.
Redacted documents
Redaction removes visible content from pages, but it does not necessarily touch document properties. A carefully redacted report can still carry its original author and title in the metadata. Metadata cleanup is a standard step after redaction.
Removing metadata
Removal makes sense when a document is leaving your control and the properties add nothing for the recipient.
Our free Remove metadata tool strips the document information dictionary (title, author, subject, keywords, creator, producer, and dates) and the XMP metadata stream, along with application-specific private data stored at the document level. It runs entirely in your browser, so the file is not uploaded.
Steps:
- Open Remove metadata and choose your PDF.
- Run the tool and download the cleaned copy.
- Open the cleaned file's document properties to confirm the fields are empty.
A note on dates: your operating system keeps its own created and modified timestamps for every file on disk, and some property dialogs show those alongside PDF fields. Those file system dates are not stored inside the PDF and do not travel with it as document metadata, although some transfer methods, such as ZIP archives, can carry file timestamps of their own.
What this tool does not remove
To stay accurate about scope, document-level metadata removal does not change the page content. It does not remove:
- Names or details that appear visibly on pages. Use Redact PDF for those.
- Metadata embedded inside individual images.
- Comments, annotations, or form field values.
- Embedded attachments.
If those matter for your document, handle them separately. For form fields, Flatten PDF turns filled values into fixed page content. For a thorough cleanup of everything, including images, one reliable approach is to redact or rasterize the relevant pages, since rasterized pages contain only pixels.
Editing metadata instead
Sometimes you do not want blank metadata. You want correct metadata. Good properties are useful:
- Search engines may use the PDF title in results when the document is published online. A title like "Document1" looks careless, while "2026 Annual Accessibility Report" is clear.
- Browsers and viewers often show the title in the tab or window.
- Document management systems index author, subject, and keywords.
- Archival workflows may require meaningful properties.
Our free Edit metadata tool lets you set the title, author, subject, keywords, creator, and producer. A common approach for public documents is to set the author to the organization rather than an individual, write a descriptive title, and add a handful of keywords separated by commas.
Remove, then edit
For documents you publish, a two-step process works well: remove all metadata first to clear the XMP history and leftover fields, then add back only the properties you want with Edit metadata. The result is a file that carries exactly the information you chose and nothing more.
Native options
Several applications offer built-in metadata controls:
- Word processors often include a "document inspector" feature that can remove personal information before you export to PDF. Cleaning the source before export prevents much of the problem.
- Full-featured PDF editors typically include options to remove hidden information or sanitize documents, often in paid tiers.
- Preview on macOS lets you view some properties but offers limited editing.
- Command-line tools are available for technical users who need to process many files.
Whichever tool you use, check the result by opening the document properties afterward.
Building a sharing checklist
If you regularly send documents outside your organization, a short routine helps:
- Remove comments and tracked changes in the source application before export.
- Export to PDF.
- Redact any sensitive content visible on the pages.
- Remove metadata, then add back a clean title if the document is public.
- Flatten forms if the recipient should not change values.
- Open the final file and check the properties and a few pages.
Key takeaways
- PDF metadata includes title, author, subject, keywords, creator, producer, dates, and XMP data.
- It is invisible on the page but easy for anyone to read in a viewer.
- Remove it when a document leaves your control and the details add nothing.
- Metadata removal does not touch page content, image metadata, comments, or attachments, so handle those separately.
- For public documents, clear the metadata and then set a clean, descriptive title.
Metadata is one of the easiest privacy leaks to fix. A few seconds with Remove metadata before you hit send is a small price for sharing only what you intended.