PDF Commons

Remove PDF metadata

Removes the document info dictionary and the catalog metadata stream. Page content stays. It runs in this tab. The file is not uploaded.

A $6 day pass is 24 hours, charged once, and it does not renew. Pricing

No ad script loads from this choice. It stays in this browser.

What this removal changes

Remove PDF metadata takes one PDF and downloads without-metadata.pdf. The original file is not rewritten. The result is still one PDF, not a zip. When the document info dictionary exists, the worker deletes Title, Author, Subject, Keywords, Creator, Producer, CreationDate, and ModDate. When the catalog Metadata entry exists, the worker deletes that entry and the catalog metadata stream. Page content stays. Visible text stays. A drawn rectangle stays. This page does not edit a content stream and does not draw anything. It does not remove font names, and it does not remove image XMP inside a picture. A reloaded document has no title and no author. Those values are absent, not empty strings. This page does not set or remove a password.

The removal runs in this tab with the open-source workers. The file is not uploaded. There is no GridFS copy. Open the network panel, drop the file, and the PDF should not appear as an upload. Closing the tab drops the copy. The form does not ask for an email. PDF to Word at /pdf-to-word uploads. OCR is /ocr-pdf. Heavy compress is /compress-pdf-heavy. Unlock is /unlock-pdf and protect is /protect-pdf.

There is no published page cap. The tab has to hold the file. There is no account on this page. The day pass is $6, charged once, for 24 hours, and it does not renew. The year is $60 for 365 days. The $6 day pass and the $60 year are for server jobs. This page does not need them. The free server file is one PDF to Word conversion, up to 20 pages. That 20-page server cap is not this tool. OCR and heavy compress are not free.

An empty file, and any file that cannot be opened, fails with "This PDF could not be read." A PDF with no pages fails with "This PDF has no pages." A locked PDF fails with "Input document to `PDFDocument.load` is encrypted. You can use `PDFDocument.load(..., { ignoreEncryption: true })` if you wish to load the document anyways." This page does not guess a password. Unlock PDF is the separate page.

Steps

  1. Drop the file. It stays in this tab.
  2. Set the one option on the page, if it has one.
  3. Download the result. Closing the tab drops the copy.

Questions

What does this removal clear?

Eight info keys, when the info dictionary exists: Title, Author, Subject, Keywords, Creator, Producer, CreationDate, and ModDate. It also removes the catalog Metadata entry and the catalog metadata stream. Page content, visible text, and font names stay. Image XMP inside a picture is not removed. The download is without-metadata.pdf. Title and author on a reloaded file are absent, not empty strings.

Does remove PDF metadata upload the file?

No. It runs in this tab. The file is not uploaded. There is no GridFS copy. Closing the tab drops the copy. The download is without-metadata.pdf. PDF to Word at /pdf-to-word is the page that uploads.

Does the page content stay?

Yes. Page content stays, including visible words and a rectangle drawn on the page. This tool does not edit a content stream. Only the document info dictionary and the catalog metadata stream are removed. Font names and image XMP are not cleared.

Do I need the day pass or the year?

No. The day pass is $6, charged once, for 24 hours, and it does not renew. The year is $60 for 365 days. Those prices are for server jobs. This page does not need them. The 20-page server cap is not this tool. There is no account here.

Also