PDF Metadata
Reads and removes title, author and producer details.
This tool runs on JavaScript. To use it, turn on JavaScript in your browser and reload the page.
THE DETAILS BEHIND THE DOCUMENT.
PDF Metadata reads the document details stored inside a PDF file. The file section holds the name, the size, the PDF version and the page count. The document details section shows the title, the author, the subject, the keywords, the program that created the document, the software that produced the PDF and the dates of creation and modification. The page size section lists the page sizes in the document in millimetres and how many pages there are of each size; that makes telling A4 (210 × 297 mm) from Letter easy.
One of the most useful lines is the text layer: the tool looks at the first pages and tells you whether there is selectable text in the document, showing the first lines if there is. If there is no text layer, the pages are most likely a scanned image; text tools such as “PDF → Word” and “PDF → Excel” give no result with documents like that. It is also used to see, before sharing a document, whether your name or your organisation's details are still inside it. Below the report a copy with its metadata removed is also ready: the document information (title, author, program, dates) and XMP are taken out, while pages, bookmarks and form fields stay as they are. For digitally signed documents no copy is made, because the signature would break.
From start to finish
- Upload your PDF. Drag the file in or select it. Documents up to 1 GB are read.
- Read the report. The file, document details, page size and content sections appear line by line.
- Download the clean copy. Use “Download file without metadata” to get the copy without author details; save the report as JSON if you like. The copies on the server are deleted when you leave the tool.
Frequently asked questions
What does “no text layer” mean?
The pages are stored as an image rather than as writing; the document has been scanned or made from photographs. In these documents the text cannot be selected or searched and the text tools do not work. That needs text recognition (OCR); the “PDF OCR (Text Recognition)” tool turns such pages into a searchable PDF or text.
Can I remove the author details?
Yes. When the report appears, press “Download file without metadata”: you get a copy with the title, author, program and date details and the XMP removed. Changing them (writing a new author) is not available. The other PDF tools of RUPO Studio, on the other hand, keep the title and the author.
What does the PDF version show?
Which PDF standard the document was written to, 1.4 or 1.7 for example. In everyday use it is usually not important; very old or very new versions can cause compatibility trouble in some viewers.
Does my file stay on the server?
No. Your file reaches the server only to be processed and is deleted when the job is done; the output is deleted too once you download it and leave the tool. You don't need an account, a name or an email address, and your files are never sent to a third-party service. It is free to use: we set no hourly or daily quota on the number of jobs, we don't make you watch an ad before processing, and we add no watermark to the output.
Other PDF tools
RUPO Studio