How to batch delete PDF author, title, subject, and keyword metadata
Translation:EnglishFrançaisDeutschEspañol日本語한êµì–´ï¼ŒUpdate Time:2026-08-18 14:28:05
Many PDF files automatically write metadata such as title, author, subject, keywords, creation time, and application during generation, editing, or conversion. If this information is inaccurate or unsuitable to retain before external distribution, archiving, or uploading to platforms, it needs to be cleaned up uniformly. This article uses HeSoft Doc Batch Tool as an example to demonstrate how to batch delete document property information such as author, title, subject, and keywords from multiple PDF files, helping users reduce repetitive operations and improve file processing efficiency.
In daily office work, archival filing, contract circulation, thesis material organization, or before sending PDF files externally, many people overlook one detail: in addition to the page content, PDF files may also store a large amount of document property information, commonly known as PDF metadata. For example, title, author, subject, keywords, creation time, modification time, generating program, and so on. A single file can be modified manually by opening the properties window in a PDF reader, but if there are dozens or hundreds of PDF files that all need to have their author, title, subject, and keywords removed, opening and clearing them one by one is very time-consuming and prone to missed processing.
This article addresses exactly this problem: using the office software " HeSoft Doc Batch Tool " to batch-remove metadata from multiple PDF files at once, making the file properties cleaner and suitable for unified processing before sharing, archiving, submitting, or uploading to a system. Below, with before-and-after comparison images and software operation screenshots, the steps for batch PDF metadata cleanup are explained in detail.
Applicable scenarios: when do you need to batch-remove PDF metadata
PDF metadata does not appear directly on the body pages, but in Adobe Acrobat, some PDF readers, document management systems, or search indexing tools, it can be viewed through "Document Properties". For both individual and enterprise users, the following scenarios are common:
First, PDF files come from complex sources. For example, when PDFs are generated from Word, docx, doc, PPT, Excel, scanning software, or online conversion tools, the files may automatically carry the creator name, company name, document title, or software name. If the files are to be sent externally, this information may not be appropriate to retain.
Second, batch organization of historical materials. During departmental archiving, training material organization, project delivery materials, e-books, or learning material merging, the titles, authors, subjects, and keywords of different PDFs may be inconsistent in format, which can affect retrieval and display in document management systems.
Third, privacy and compliance requirements. Some PDF metadata may contain personal names, internal project codenames, generation paths, software information, and other content. Batch-removing PDF author information, title information, and keyword information before external release can reduce the risk of inadvertent disclosure.
Fourth, file standardization. Many organizations want PDF file properties to remain blank or follow unified rules before uploading to archives, repositories, or business platforms. Compared with manual processing one by one, batch-removing PDF metadata is more suitable for repetitive file tasks in office scenarios.
Effect preview: comparison of PDF properties before and after processing
Before processing: the PDF document properties retain the title, author, subject, and keywords
From the before-processing screenshot, it can be seen that after opening the PDF's "Document Properties" window, in the "Description" tab, the title field shows "Class Notes Packet", the author field shows "Ian Stone", the subject field shows "2026 Class Notes Activity", and the keywords field also contains a string of keyword content. Although this information is not directly displayed in the body of the PDF page, it is still part of the file properties.

If there is only one PDF file, manually deleting these fields may be acceptable; but if a folder contains many PDFs with similar properties, manually opening, clearing, and saving each one is very inefficient. More importantly, manual operation requires repeating the same actions over and over, making it easy to forget to clean a file or miss clearing a field.
After processing: the title, author, subject, and keyword fields have been cleared
After processing, when viewing the "Document Properties" of the same PDF file again, it can be seen that the input boxes corresponding to title, author, subject, and keywords are now empty. The screenshot also shows that the PDF file's save location has changed, indicating that the processed file was output to a new directory, making it easy to distinguish from the original file.

This kind of effect is exactly the core goal of batch-removing PDF metadata: not changing the user-visible PDF page content, but cleaning the descriptive information in the document properties, so that the files are more standardized during distribution, submission, and archiving.
Operation steps: using HeSoft Doc Batch Tool to remove PDF metadata
Step 1: Enter the PDF tools and select "Remove metadata from PDF"
After opening HeSoft Doc Batch Tool , on the left side you can see the tool categories, including Home, Task Flow, All Tools, File Name, Folder Name, File Organization, Word Tools, Excel Tools, PowerPoint Tools, PDF Tools, and so on. Since the objects to be processed this time are PDF files, first select "PDF Tools" on the left side.
In the PDF tool list, find "32. Remove metadata from PDF". In the screenshot, this function card is highlighted, and the function description indicates that it batch-removes document metadata such as title, author, time, and XMP from PDF files. When you hover over the function card, related prompts are also shown, explaining that this function is specifically used to clean PDF document property information.

The purpose of this step is to enter the correct batch processing function. For users who want to remove PDF author, title, subject, and keywords, there is no need to enter "Modify PDF metadata" or other conversion tools; instead, select the dedicated "Remove metadata from PDF" function, which can complete the clearing operation more directly.
Step 2: Add the PDF files to be processed
After entering the "Remove metadata from PDF" function page, the top of the software interface displays the current task name and shows the processing flow in a step bar: select the records to process, set the save location, and start processing. The current screenshot stays at step 1, that is, selecting the files to process.
In the upper right area of the interface, you can see buttons such as "Add Files", "Import Files from Folder", "Clear", and "More". If you only need to process a few PDFs, you can click "Add Files" to select specified files; if you need to process all PDFs in an entire directory at once, you can use "Import Files from Folder" to import the files in the folder. In the screenshot, the list has already imported 4 PDF files, named 1.pdf, 2.pdf, 3.pdf, and 4.pdf, all located in the D:\test\ directory.

The list area displays information such as sequence number, name, path, extension, creation time, and modification time. Users can confirm here whether the files were added correctly. If a file does not need to be processed, it can be removed through the delete icon in the operation column; if the import was wrong, you can also click "Clear" to select again. The bottom shows "Record count: 4", which is used to confirm the number of files for this batch process.
Step 3: Confirm the file list and proceed to the next step
After the files are imported, it is recommended to first check three key points: whether the file extensions are all pdf, whether the paths are the expected directory, and whether the file count is correct. The advantage of a batch processing tool is that it processes multiple files at once, but precisely because it is a batch operation, confirming the list before starting is very important.
After confirming that everything is correct, click the "Next" button at the bottom of the page. The software will enter step 2, "Set save location". From the post-processing screenshot, it can be seen that the processed PDFs are saved to an output directory rather than being displayed directly in their original location. Therefore, when setting the save location, it is recommended to choose an easily identifiable output folder, such as an output directory on the desktop or a dedicated results directory. The advantage of this is that the original PDF files can be retained for easy comparison, review, and rollback.
Step 4: Set the save location and start processing
After entering the save location settings, follow the interface wizard to select the output location for the processed PDFs. Since the step bar in the screenshot shows step 3 as "Start processing", after completing the save location settings, simply continue to the start processing stage. The software will process the PDFs one by one according to the file list, clean up the metadata in the document properties, and output new PDF files.
After processing is complete, it is recommended to randomly open one of the result files and use a PDF reader or Adobe Acrobat's document properties window to view the "Description" tab. If the title, author, subject, keywords, and other fields are empty, it indicates that the operation to remove PDF metadata was successful. From the effect image, it can be seen that the relevant fields in the processed 2.pdf have been cleared, achieving the purpose of batch cleanup.
Frequently asked questions and precautions
Will removing metadata change the PDF page content?
Based on the effect shown in the screenshots, this function processes PDF document property information, such as title, author, subject, keywords, and other metadata, rather than deleting body text, images, tables, or annotation content. Therefore, under normal circumstances, the PDF page content visible to the user will not change because these property fields are cleaned. However, before formally batch-processing important files, it is still recommended to test with a small number of samples first to confirm that the results meet expectations.
Why is it recommended to output to a new folder?
The most important thing in batch file processing is controllability. Saving the processed results to a new folder can avoid overwriting the original files and facilitate later verification. If it is found that a certain file should not have been cleaned, or if the original property information needs to be retained, you can still go back to the original file and process it again. The location shown in the post-processing screenshot is the output directory, and this approach is more suitable for office batch tasks.
Can it only process title, author, subject, and keywords?
The document properties window in the screenshot mainly highlights the title, author, subject, and keyword fields; the software function prompt also mentions that it can batch-remove document metadata such as title, author, time, and XMP. In actual processing, the current version of the software interface and processing results should prevail. For most office scenarios, clearing these common PDF property fields is already sufficient to meet the needs of external distribution, archiving, and standardized processing.
Do I need to back up before processing?
Backup is recommended. Although this type of operation mainly targets metadata and does not directly edit page body content, once a batch task is executed, it affects multiple files. Keeping the original files in a separate directory and outputting the processing results to a new directory is a more prudent habit for office file processing.
Summary
Batch-removing PDF author, title, subject, keywords, and other metadata may seem like a small need, but it is very practical in the processes of archival filing, external file distribution, privacy protection, and document standardization. HeSoft Doc Batch Tool , as batch file processing software designed for office scenarios, can help users turn the repetitive labor of opening PDFs one by one and clearing properties item by item into a batch process of importing files, setting the save location, and starting processing.
If you have a large number of PDF files that need document property cleanup, it is recommended to first prepare a test folder, follow the steps in this article to enter "Remove metadata from PDF" under "PDF Tools", import the files, and output them to a new directory. After confirming that the processing results are correct, then batch-process the formal folder. This can improve efficiency while also making the file processing workflow more reliable.
Keyword:Batch delete PDF metadata , PDF author title cleanup , delete PDF keywords , PDF property batch processing