From PDF reports straight into Excel.
Replacing a fragile copy-and-paste PDF macro with a direct extraction workflow.
01PDF report
02Extract & normalize
03Excel summary
What needed to change
The original macro depended on opening PDFs in Acrobat, copying text and parsing the pasted results. Number-format differences made the process fragile.
What I built
A Python pdfplumber reader extracts the PDF content directly, recognizes European and North American numeric formats, and writes mapped values into an Excel summary. VBA provides the workbook entry point.
What changed
The project notes record a sample verification in which all 17 checked values matched the source PDF. This is a specific sample check, not a universal accuracy claim. No client PDF or financial record is published here.