How to batch delete PDF author, title, subject, and keyword metadata to avoid file information leakage


Translation:EnglishFrançaisDeutschEspañol日本語한국어,Update Time:2026-08-18 14:30:21
Content note: This article reflects the software version available when it was published. Interfaces and features may change with updates; please refer to the current software. If you find an error, please let us know.
Software featured in this articleHeSoft Doc Batch Tool
Free DownloadSoftware IntroductionUser Guide

Many PDF files retain document property information such as title, author, subject, keywords, and creating application before being circulated, archived, or sent externally. If you open them one by one with a PDF reader and manually clear them, it is not only time-consuming but also easy to miss. This article explains how to use HeSoft Doc Batch Tool to import multiple PDF files at once in an office scenario, batch delete the author, title, subject, keywords, and related metadata in the PDFs, and illustrate the cleaning effect through before-and-after screenshots.

In daily office work, PDF files are often used for contracts, courseware, reports, bidding materials, training materials, archived documents, and other scenarios. Many people only focus on whether the PDF body content is correct, but overlook that the file properties may still contain metadata such as author, title, subject, keywords, application, creation time, modification time, and XMP. For files used only internally, this information may not be a major issue; however, if a PDF needs to be sent to clients, uploaded to a platform, submitted for review, or made public, residual personal names, project names, internal keywords, and generating software information in the metadata may pose privacy and compliance risks.

If you only need to handle one or two PDFs, opening tools such as Adobe Acrobat, going to document properties, manually clearing the fields, and saving can still be acceptable. But when a folder contains dozens or hundreds of PDFs, opening them one by one, checking properties, deleting title and author information, and saving and closing each one consumes a great deal of time and is also prone to omissions. This article aims to solve this typical repetitive task: using the "Delete Metadata in PDF" feature in the office software " HeSoft Doc Batch Tool " to batch delete metadata such as author, title, subject, and keywords from many PDF files.

As can be seen from the screenshot, the tool is positioned as batch document processing software, with functions organized on the left into categories such as PDF tools, Word tools, Excel tools, PowerPoint tools, and image tools. This article focuses on demonstrating the PDF metadata cleanup process, which is suitable for office users who need to uniformly clean PDF property information, reduce repetitive operations, and improve file delivery efficiency.

Applicable scenarios: when batch deletion of PDF metadata is needed

PDF metadata does not appear directly on the body pages, but it can be viewed in the document properties. The before-processing effect shown in the screenshot indicates that a PDF file's properties include fields such as title, author, subject, and keywords. For example, the title is Class Notes Packet, the author is Ian Stone, the subject is 2026 Class Notes Activity, and the keywords include learning, study, pdf, simulated metadata, 2026, and other content. This information may come from the document generation process or may be automatically written by the PDF production program.

Batch cleaning of PDF metadata is especially valuable in the following scenarios:

  • Before sending materials externally: For example, quotations, project descriptions, course materials, and client reports need to avoid exposing internal author names, project codenames, editing software, or historical information.
  • Before archiving files: When a company uniformly archives PDFs, it is desirable to keep document properties clean and reduce interference from irrelevant fields in retrieval and management.
  • Batch organizing historical PDFs: PDFs collected from multiple sources may carry different authors, titles, subjects, and keywords. After unified cleanup, subsequent naming, classification, and distribution become easier.
  • Privacy compliance checks: Before uploading to a platform, submitting materials, or publishing publicly, delete hidden information such as author, title, and keywords in PDFs to reduce the risk of information leakage.
  • Cleanup after batch conversion: PDFs exported in batches from Word, doc, docx, PPT, Excel, or other systems may retain source file properties. Cleaning metadata can make the final PDFs more standardized.

It should be noted that this article discusses metadata cleanup in PDF document properties, not deleting page body content, annotations, bookmarks, or visible watermarks. In other words, after cleanup, the PDF page content remains intact; only fields such as title, author, subject, and keywords in the properties panel are deleted or cleared.

Effect preview: before processing, PDF properties contain author, title, subject, and keywords

Below is a screenshot of the PDF document properties before processing. It can be seen that on the "Description" tab, the title, author, subject, and keyword fields all have content. The area marked with a red box indicates the main metadata fields that need to be batch deleted.

image-Batch delete PDF metadata,PDF author and title cleanup,PDF keyword deletion

From this image, it can be seen that PDF metadata usually does not appear directly in the page body, but it can be viewed as long as the document properties are opened. For externally distributed files, the author name, course or project subject, keywords, generating program, and other information may not be necessary for the recipient to know. If these fields need to be cleared one by one across a large number of PDFs, the operation becomes very tedious.

Effect preview: after processing, PDF title, author, subject, and keywords have been cleared

Below is the effect after processing with HeSoft Doc Batch Tool . When the document properties of the same PDF are opened again, it can be seen that the title, author, subject, and keyword fields are already empty. The screenshot also shows that the file location has changed to the output directory, indicating that the software generated the processed PDF file instead of requiring the user to repeatedly modify the original file manually.

image-Batch delete PDF metadata,PDF author and title cleanup,PDF keyword deletion

The change after processing is very intuitive: the identifiable text information in the document properties has been cleaned up, while the PDF page content remains intact. For office scenarios requiring batch deletion of PDF author, title, subject, and keywords, this method is more stable than manual processing one by one and also makes it easier to check results in a unified manner.

Operation steps: using HeSoft Doc Batch Tool to batch delete PDF metadata

The specific operations are explained below according to the order of the screenshots. Different versions may have slight differences in the interface, but the core process is usually: enter the PDF tools category, select "Delete Metadata in PDF", import PDF files, proceed to the next step to set the save location, and finally start processing.

Step 1: Enter the PDF tools category and find the delete PDF metadata function

After opening HeSoft Doc Batch Tool , select "PDF Tools" from the function categories on the left. In the upper-left corner of the screenshot, the software name is shown as " HeSoft Doc Batch Tool ", and the current version is displayed as v1.25.1. The main area lists multiple PDF-related batch processing functions in card form, such as deleting blank pages in PDFs, PDF optimization and compression, repairing PDFs, modifying PDF metadata, deleting bookmarks in PDFs, converting PDFs to Word, and converting PDFs to PowerPoint.

In this task, the function to select is number 32, "Delete Metadata in PDF". The mouse hover prompt in the screenshot shows that this function is used to batch delete document metadata such as title, author, time, and XMP in PDF files. This is consistent with the goal of this article: batch cleaning PDF property information, especially fields such as author, title, subject, and keywords.

image-Batch delete PDF metadata,PDF author and title cleanup,PDF keyword deletion

The purpose of this step is to find the correct batch processing entry in the software. After selecting "Delete Metadata in PDF", the software enters a dedicated task page, and then multiple PDF files can be added and processed at once.

Step 2: Add the PDF files to be processed

After entering the "Delete Metadata in PDF" page, the top of the interface displays the current function name, and there is a "Return to main panel" button in the upper-left corner for returning to the tool list. In the upper-right area, buttons such as "Add Files", "Import Files from Folder", "Clear", and "More" can be seen. Instead of opening each file one by one with a PDF reader, directly import the PDFs to be processed into the list in batches.

If only a few specified PDFs need to be processed, click "Add Files"; if a folder contains many PDFs that need unified cleanup, use "Import Files from Folder". In the screenshot, four PDF files have already been imported, namely 1.pdf, 2.pdf, 3.pdf, and 4.pdf, located in the D:\test\ directory. The list also displays information such as extension, creation time, and modification time, making it easy to confirm whether the imported files are correct.

image-Batch delete PDF metadata,PDF author and title cleanup,PDF keyword deletion

The purpose of this step is to gather all PDFs whose metadata needs to be cleaned into the same processing task. The expected result is that the file list shows the PDF records to be processed, and the summary area at the bottom shows the record count. In the screenshot, "Record count: 4" indicates that the current task will process four PDF files.

Step 3: Check the file list and remove incorrect records if necessary

Before batch processing, it is recommended to check the file names and paths first. Because deleting PDF metadata is a batch operation that modifies document properties, if a PDF that should not be processed is imported by mistake, although the page content is usually not deleted, the file properties will be cleaned, and the information may need to be re-entered later.

In the screenshot, there is an "Actions" column on the right side of the list, and each record has a delete icon next to it. If a certain PDF should not participate in this processing, it can be removed from the list through that action. The "Clear" button at the top is suitable for directly clearing the current list when there are many import errors, and then adding files again.

The purpose of this step is to reduce batch processing errors. The expected result is that the list contains only PDF files whose metadata such as title, author, subject, and keywords truly need to be deleted.

Step 4: Click Next and continue with the wizard to set the save location

The progress bar at the top of the interface shows three stages: select records to process, set save location, and start processing. The current screenshot is at step 1, and there is a "Next" button at the bottom of the page. After confirming that the file list is correct, click "Next" to enter the save location settings.

Although the screenshot does not show the detailed controls on the save location page, it can be reasonably inferred from the progress bar that the software will require specifying a save location for the processed PDFs. It is recommended to choose a separate output directory instead of directly overwriting the original files. This has two advantages: first, the original PDFs can be kept as backups; second, after processing is complete, it is easier to check the output files in one place. The after-processing effect image shows that the output path is in the hesoft-output directory on the user's desktop, which also indicates that the software saves the processed files to the output location.

The purpose of this step is to clarify where the cleaned PDFs will be saved. The expected result is that the software knows the output directory and is ready to enter the final batch processing stage.

Step 5: Start processing and check the results

After completing the save location settings, follow the wizard to the "Start Processing" stage. Once processing starts, the software will execute the metadata deletion task on each PDF in the list one by one. For users, there is no need to repeatedly open each PDF or manually enter document properties to clear each field.

After processing is complete, it is recommended to randomly open several output PDFs and use a PDF reader to check "Document Properties" or a similar information panel to confirm whether fields such as title, author, subject, and keywords are empty. The after-processing screenshot in this article shows that these fields have been cleared, the modification date has been updated, and the output location has also changed, indicating that the batch cleanup task has taken effect.

The purpose of this step is to complete the batch deletion of PDF metadata and verify the results. The expected result is that attribute information such as author, title, subject, and keywords in multiple PDF files has been uniformly cleaned, and the user obtains a batch of PDF files suitable for external distribution or archiving.

Frequently asked questions and precautions

1. Will deleting PDF metadata delete the body content?

As can be seen from the before and after screenshots, what is cleaned are fields such as title, author, subject, and keywords in the PDF document properties, not the body content on the PDF pages. Under normal circumstances, deleting metadata does not affect page text, images, or layout. However, before formally processing important files, it is still recommended to test with a small number of samples and keep backups of the original files.

2. What is the difference between PDF metadata and file name?

The file name is the name seen in the operating system, such as 2.pdf; metadata is the internal PDF properties, such as title, author, subject, and keywords. Even if the file name appears to contain no sensitive information, the PDF properties may still retain an author name or internal project keywords. Therefore, before sending files externally, changing only the file name is not enough; it is best to check and clean the PDF metadata at the same time.

3. Why batch process instead of modifying each file one by one with a PDF reader?

A single PDF can be modified manually, but a large number of PDFs creates repetitive work. The value of batch processing software lies in applying the same rules to multiple files: import once, output uniformly, and check centrally. For office staff, data administrators, legal personnel, and training document producers, this can significantly reduce mechanical operation time and lower the risk of missed deletions.

4. Can PDFs from multiple folders be processed at the same time?

The screenshot shows two entry points: "Add Files" and "Import Files from Folder". In actual operation, the import method can be selected according to how the files are organized. If PDFs are scattered across multiple directories, files can be added in batches; if they are concentrated in one folder, importing from the folder is more efficient. Be sure to check the paths and record count in the list before processing.

5. What checks still need to be done after processing?

It is recommended to perform at least two types of checks: first, open the output PDF to confirm that the page content is normal; second, view the document properties to confirm that fields such as title, author, subject, and keywords have been cleared. For especially important external files, spot-check PDFs from different sources, with different page counts, and generated by different programs to ensure the results meet expectations.

Summary: use batch processing to reduce the repetitive work of PDF metadata cleanup

Batch deleting PDF author, title, subject, and keyword metadata is a seemingly small but very practical office need. It relates to the professionalism of external file distribution, privacy protection, and standards for document management. If it relies on manually opening PDFs one by one to modify properties, it is not only inefficient but also prone to omissions.

HeSoft Doc Batch Tool turns this type of repetitive operation into a clear PDF batch processing function. Users only need to enter "PDF Tools", select "Delete Metadata in PDF", import multiple PDF files, set the save location according to the wizard, and start processing to batch clean document metadata such as title, author, time, and XMP in PDFs. For teams that frequently handle office documents such as PDF, Word, docx, Excel, and PPT, this kind of batch processing tool can significantly reduce manual operations and leave more time for reviewing content and delivering results.

If you are about to send a batch of PDFs to clients, upload them to a platform, or archive them, it is recommended to first perform a PDF metadata cleanup according to the steps in this article and conduct spot checks on the output files. This can both improve processing efficiency and make file delivery more secure and standardized.


Keyword:Batch delete PDF metadata , PDF author and title cleanup , PDF keyword deletion
Creation Time:2026-08-18 14:30:08
Software featured in this articleHeSoft Doc Batch Tool
Free DownloadSoftware IntroductionUser Guide
Content note: This article reflects the software version available when it was published. Interfaces and features may change with updates; please refer to the current software. If you find an error, please let us know.