Skip to content

File metadata remover

Take the hidden data out of a file before you send it.

It excises C2PA Content Credentials, EXIF, XMP, IPTC, and PDF or Office document properties from the file, and shows you exactly what it took out.

How to use it

Choose a PDF, PNG, JPEG or Office file. We show you what it carries, you pick what to take out, and we return the file with those parts excised. The picture and the text are untouched, because nothing is re-saved through a converter.

When it helps

A photo can carry the exact spot it was taken and the phone that took it. An Office file carries the author and the company that last saved it. A PDF carries the title and author fields. None of it shows when you open the file, and all of it travels with the file when you share it. Some of it is also wrong about your own work: a picture you took can carry a record naming an app it passed through, and some sites read those records when they decide how to describe a post.

Common questions

Which files can I clean?

PDF, PNG, JPEG, and Office files such as .docx, .xlsx, .pptx and their OpenDocument equivalents.

Does this change the picture or the text?

No. We cut the metadata out of the container and leave the rest of the bytes as they are. For images the picture data is untouched. For a PDF the pages, text, bookmarks and form fields all survive. The file is rewritten as it is saved, so the bytes differ while the content does not.

What can this tool not take out?

Two things, and neither of them is in the file. Some image generators work a faint pattern into the picture data itself, which is part of the image rather than a field attached to it, so a tool that deletes fields does not reach it and we make no claim about it. And some Content Credentials register a fingerprint worked out from the picture and keep it in a separate database, with nothing about it written into your file, so there is nothing in the file to take out. When we find one we say so before you run and again on the result. What we do take out is the data written into the file: the Content Credentials stored in it, EXIF, XMP, IPTC, and PDF or Office document properties.

What does a cleaned PDF still contain?

A PDF saved by our engine names the engine itself in its producer and creator fields, and carries an XMP packet describing that. We cannot prevent it, because the library writes those back after the removal has run. What is left names our software, not you. The title, author, subject and keywords you filled in are gone, along with page-level XMP.

My file is signed. What happens?

We tell you before anything runs, and we stop. The signature covers the metadata, so taking the metadata out breaks it. You decide which one you want; we do not make that choice quietly.

Do you keep my file?

The cleaned file is held only long enough for you to download it, then it is deleted. What the file carried is shown to you and is not logged or shared.

Related tools