#1 ·
So, I find myself facing a bit of a logistical headache and am reaching out to see if anyone here possesses the technical wizardry required to solve this...
I am currently staring down a mountain of PDF documents—text-heavy files, mind you—where certain sections absolutely must be redacted to ensure privacy. To be specific, we are dealing with lists containing full names paired with Social Security numbers. My situation is such that I have a list of individuals whose names must remain visible, while I also possess a separate roster of the specific people who need to be completely scrubbed from the records.
While I am well aware that one could manually redact these entries using various PDF editors equipped with a "redact" tool, doing so one by one across a massive volume of files is simply not a viable use of my time. Is there any way to automate this process for a large batch of documents?
To illustrate the dilemma, suppose I have a list within a PDF that looks like this:
1. Anne Aniston, SSN 000-00-0000
2. Mark Miller, SSN 000-00-0000
3. Isaac Ives, SSN 000-00-0000
Now, imagine that this list, along with various other data points, is scattered across dozens of different PDFs. In every single one of them, I need to black out Mark Miller and Isaac Ives, leaving only Anne Aniston's information intact. Does anyone know of a way to automate this so I can actually get through this workload without losing my mind...
Any idea?
I am currently staring down a mountain of PDF documents—text-heavy files, mind you—where certain sections absolutely must be redacted to ensure privacy. To be specific, we are dealing with lists containing full names paired with Social Security numbers. My situation is such that I have a list of individuals whose names must remain visible, while I also possess a separate roster of the specific people who need to be completely scrubbed from the records.
While I am well aware that one could manually redact these entries using various PDF editors equipped with a "redact" tool, doing so one by one across a massive volume of files is simply not a viable use of my time. Is there any way to automate this process for a large batch of documents?
To illustrate the dilemma, suppose I have a list within a PDF that looks like this:
1. Anne Aniston, SSN 000-00-0000
2. Mark Miller, SSN 000-00-0000
3. Isaac Ives, SSN 000-00-0000
Now, imagine that this list, along with various other data points, is scattered across dozens of different PDFs. In every single one of them, I need to black out Mark Miller and Isaac Ives, leaving only Anne Aniston's information intact. Does anyone know of a way to automate this so I can actually get through this workload without losing my mind...
Any idea?