Common mistakes when manually redacting a PDF
The most common way to «protect» a PDF is to open it in an editor, draw a black rectangle over the name or ID, and save the file. The problem: this does not anonymize anything. The text remains in the PDF's underlying layer, accessible with a click, copy-paste, or extraction tools. These are the most common mistakes—and their real consequences.
The 5 most frequent mistakes
1. Visual-only hiding
The black rectangle covers the data on screen, but the underlying text remains selectable and searchable.
2. Forgetting metadata
The PDF author, title, and properties may contain names and identifying data.
3. Leaving comments and annotations
Word/PDF review layers may contain personal data in hidden notes.
4. Partial anonymization
The name is hidden but not the ID, phone, address, or indirect references in the text.
5. Not verifying the result
Sending the document without checking that data is not recoverable.
Real consequences
These mistakes are not theoretical. In 2024, a European public body published documents with «hidden» personal data that were recovered in seconds by a journalist using the PDF search function. The result: GDPR fine, public retraction, and lasting reputational damage.
How to anonymize correctly
- Use tools that remove content, not just hide it visually.
- Clean metadata and hidden layers before sharing the document.
- Verify the result by trying to search and select supposedly removed data.
- For high volumes, automate with AI to avoid human errors.
Want to stop manually redacting PDFs and anonymize for real?
Try PDF anonymization