Every PDF you create carries invisible metadata - Your name, company, software version, edit history, and sometimes GPS coordinates from embedded photos. Stripping this data before sharing legal, financial, or tender documents is essential and takes less than a minute with the right tools.
- PDF metadata includes author, organisation, software, revision count, and creation timestamps - All invisible to casual readers but trivially extractable.
- Images embedded in a PDF can carry EXIF data including GPS location, camera model, and shooting date.
- Metadata leaks have compromised legal negotiations, competitive tenders, and anonymous whistleblower filings.
- PDFBEAR's Redact PDF and Flatten PDF tools permanently strip metadata before you download the cleaned file.
That bid proposal or court filing you are about to send carries your firm's name, edit count, and timestamps - Unless you strip the metadata first.
What Metadata Is Hidden Inside Your PDFs

When your word processor or design application exports a file to PDF it embeds a surprisingly detailed dossier about the document and the person who created it. This information travels invisibly with every copy of the file you share. Recipients with even basic tools - Or a quick look at File > Properties in Adobe Acrobat - Can read all of it.
The standard metadata fields stored in the PDF's XMP (Extensible Metadata Platform) packet and Document Information Dictionary include:
| Metadata field | Typical value | Why it matters |
|---|---|---|
| Author | Full name of the person who created the file | Reveals the drafter's identity even in "anonymous" submissions |
| Creator / Producer | Microsoft Word 365, Adobe InDesign 18, LibreOffice 7.x | Exposes your software stack; old versions signal unpatched vulnerabilities |
| Organization / Company | Acme Legal LLP | Reveals the originating firm before negotiations are complete |
| Creation date | 2026-06-10T09:14:22Z | Shows when drafting started - Useful intelligence in litigation |
| Modification date | 2026-06-24T17:53:01Z | Reveals how recently the document was touched |
| Revision number | 14 | Indicates the document went through 13 prior drafts - Implying negotiation or uncertainty |
| Document title | DRAFT_Settlement_Offer_v14_FINAL.docx | The original filename can expose intent or internal codenames |
| Keywords / Subject | Confidential, Privileged | Can contradict the document's apparent purpose |
EXIF Data in Embedded Images: The GPS Problem
Many professionals embed photographs, scanned pages, or product images directly into PDF documents. What they rarely realise is that those images carry their own layer of metadata called EXIF (Exchangeable Image File Format) data, and that data survives the PDF export process intact.
EXIF fields commonly embedded in PDF images include GPS latitude and longitude (pinpointing where the photo was taken), camera make and model, lens focal length, shooting date and time, and even the smartphone's iOS or Android version. For a document photographed at a confidential client site, this could reveal the meeting location. For a whistleblower submitting evidence, it could identify where they are based.
Stripping metadata from a PDF therefore requires addressing at least two distinct data stores: the document-level XMP/DocInfo fields and the image-level EXIF data inside each embedded image object. A tool that only removes document properties while leaving image EXIF intact is not fully cleaning the file.
Real-World Scenarios Where Metadata Has Caused Harm
These are not edge cases. Metadata leaks occur constantly across professional sectors:
Competitive tendering: A construction company submits a bid as a PDF. The document's revision history shows 22 drafts and the "Creator" field names the estimating software's licensing company - Giving the tender committee and competing bidders insight into the firm's internal deliberation process. In some jurisdictions, procurement regulations explicitly require metadata-clean submissions.
Legal negotiations: A law firm sends a settlement proposal. The title embedded in the PDF metadata reads "DRAFT_Settlement_Acme_LOW_OFFER_v3.docx" - Inadvertently revealing the client's negotiating floor before mediation begins.
Court filings: Counsel files a motion with a PDF that retains the original author name. The opposing party's e-discovery tools extract this and use it to challenge the authenticity of the document's stated authorship.
Whistleblowing: A journalist or regulator receives a PDF sent by an anonymous source. The document's author field contains the source's full Windows login name, and an embedded JPEG carries GPS coordinates pointing to the office building where the photo was taken. Anonymity is destroyed.
Property transactions: A due-diligence report is shared by email. The embedded revision count (47 revisions) and creation date (six weeks before the transaction was announced) reveal the timeline of internal discussions that the seller wanted to keep private.
What Redact PDF and Flatten PDF Remove
PDFBEAR provides two complementary tools for cleaning PDF metadata. Understanding what each one does helps you choose the right approach for your situation.
For most users, running Redact PDF is the primary step - It clears document-level metadata and strips EXIF from images. Following it up with Flatten PDF collapses annotation layers (including comments your colleagues left during review), eliminates interactive form fields, and merges everything into a static, archival-grade PDF. The combination gives you the cleanest possible output.
How to Remove PDF Metadata With PDFBEAR
The workflow takes under two minutes even for large documents:
- Navigate to Redact PDF. Upload your document by drag-and-drop or file browser. Files up to 50 MB are accepted on the free tier.
- If you have specific text or image regions to redact, mark them now. If you only want metadata removed without visual redaction, you can proceed without marking any regions - The tool will still strip all document metadata during processing.
- Click Apply Redaction and download the cleaned file.
- For a second pass, upload the downloaded file to Flatten PDF. This merges annotation layers, comments, and form fields into the base content layer so they cannot be reconstructed.
- Verify the result: open the cleaned PDF in Acrobat or your viewer of choice, go to File > Properties > Description and confirm the Author, Subject, Keywords, and Creator fields are blank.
All processing happens over HTTPS-encrypted connections. Your files are not reviewed by any human, and free-tier files are automatically purged within 14 days of last activity. No software installation is needed and no watermarks are added to the output.
What Stays in the File After Cleaning
A common question is: what does the PDF retain after full metadata removal? The answer matters for usability.
| Element | Retained after Redact + Flatten? | Notes |
|---|---|---|
| Visible text content | Yes | All body text, headings, and captions are preserved |
| Images (visual content) | Yes | Images remain; only their EXIF sidecar data is removed |
| Fonts and typography | Yes | Embedded font subsets are retained so the layout is unchanged |
| Hyperlinks | Yes (after Redact alone); collapsed after Flatten | If you need clickable links in the output, do not Flatten |
| Digital signatures | Signature is invalidated by any edit | Re-sign the document after cleaning if a valid signature is required |
| PDF/A archival compliance | Preserved if the input was PDF/A | Processing maintains structural integrity |
One practical note: if the document carries a digital signature, any modification - Including metadata stripping - Will break the cryptographic validity of that signature. The document will show as "signature invalid" in Acrobat. For signed documents, strip metadata before the signing step, or re-sign after cleaning. You can use the full suite of PDFBEAR tools - Including Compress PDF to reduce file size after cleaning - As part of your pre-signing workflow.
Compare PDF tools