PDF metadata can include a title, author, subject, keywords, dates, creator application, producer, and XMP records. Removing those properties reduces unintended disclosure, but it does not remove visible text, annotations, attachments, or every possible embedded object.
What is PDF metadata?
Metadata describes the document rather than serving as its main page content. Adobe distinguishes basic document properties—such as title, author, subject, and keywords—from broader metadata that can include creation and modification dates, extended schemas, and application-specific information.
PDFs can also contain XMP, an XML-based metadata system embedded in the file. It exists to help applications catalogue, search, classify, and manage documents, but those same descriptive fields may reveal information the sender did not intend to distribute.
| Information | What it may reveal | Where to review |
|---|---|---|
| Author | A person's name, username, or organization | Document properties |
| Creator / producer | The software or workflow used to create the PDF | Document properties |
| Title, subject, keywords | Internal labels, project names, or classification terms | Description and additional metadata |
| Creation / modification dates | Timing and document-history clues | Additional metadata or XMP |
| Custom XMP fields | Application-, industry-, or workflow-specific data | Advanced metadata view |
Metadata is not the same as page content
Changing an author field does not alter a name printed on the page. Conversely, visually covering text with a shape is not necessarily secure redaction: the underlying text may remain selectable, searchable, or extractable.
Use metadata cleaning for document properties. Use a dedicated redaction workflow for confidential page content, and inspect the final exported file rather than assuming a visual edit removed the underlying data.
What the FreNiMi cleaner targets
The FreNiMi PDF Metadata Remover clears the standard document information dictionary and removes the catalog's XMP metadata reference. It targets commonly exposed values such as title, author, subject, keywords, creator, and producer.
It does not claim to sanitize every PDF object. A complex PDF may contain:
- Annotations, comments, or review history
- Embedded files and portfolio attachments
- Form field values or JavaScript
- Layers, hidden objects, or alternate content
- Visible names, signatures, account numbers, or images
A safer sharing workflow
Work from a copy
Preserve the original for records and create a separate distribution version.
Review visible content
Check every page, annotation, form field, and attachment before considering metadata.
Clean standard metadata
Inspect the detected properties, remove them, and download the separate cleaned copy.
Reopen and verify
Review document properties and search the exported file for information that should not remain.
Related PDF privacy tools
Frequently asked questions
Can a PDF show who created it?
Yes. Standard document properties can contain an author name, while creator and producer fields may identify the application that created or exported the PDF.
Does removing PDF metadata redact visible information?
No. Metadata cleaning does not remove names, account numbers, images, or other information visibly printed on a page.
Can a cleaned PDF still contain hidden information?
Yes. Attachments, annotations, form values, scripts, layers, and other embedded objects are separate from standard document properties and require their own review.