Story Beyond the Eye: Glyph Positions Break PDF Text Redaction

Bland, Maxwell; Iyer, Anushya; Levchenko, Kirill

Computer Science > Cryptography and Security

arXiv:2206.02285 (cs)

[Submitted on 5 Jun 2022 (v1), last revised 14 Nov 2022 (this version, v3)]

Title:Story Beyond the Eye: Glyph Positions Break PDF Text Redaction

Authors:Maxwell Bland, Anushya Iyer, Kirill Levchenko

View PDF

Abstract:In this work we find that many current redactions of PDF text are insecure due to non-redacted character positioning information. In particular, subpixel-sized horizontal shifts in redacted and non-redacted characters can be recovered and used to effectively deredact first and last names. Unfortunately these findings affect redactions where the text underneath the black box is removed from the PDF.
We demonstrate these findings by performing a comprehensive vulnerability assessment of common PDF redaction types. We examine 11 popular PDF redaction tools, including Adobe Acrobat, and find that they leak information about redacted text. We also effectively deredact hundreds of real-world PDF redactions, including those found in OIG investigation reports and FOIA responses.
To correct the problem, we have released open source algorithms to fix trivial redactions and reduce the amount of information leaked by nonexcising redactions (where the text underneath the redaction is copy-pastable). We have also notified the developers of the studied redaction tools. We have notified the Office of Inspector General, the Free Law Project, PACER, Adobe, Microsoft, and the US Department of Justice. We are working with several of these groups to prevent our discoveries from being used for malicious purposes.

Subjects:	Cryptography and Security (cs.CR)
Cite as:	arXiv:2206.02285 [cs.CR]
	(or arXiv:2206.02285v3 [cs.CR] for this version)
	https://doi.org/10.48550/arXiv.2206.02285

Submission history

From: Maxwell Bland [view email]
[v1] Sun, 5 Jun 2022 23:10:23 UTC (553 KB)
[v2] Mon, 15 Aug 2022 17:46:11 UTC (641 KB)
[v3] Mon, 14 Nov 2022 01:09:17 UTC (653 KB)

Computer Science > Cryptography and Security

Title:Story Beyond the Eye: Glyph Positions Break PDF Text Redaction

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Cryptography and Security

Title:Story Beyond the Eye: Glyph Positions Break PDF Text Redaction

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators