Complex PDF Tables to Excel

Employer not named by the sourceRemote

Full Stack

Apply on the company’s site

Frontier is not the employer and does not collect applications.

About this role

Python, Data Processing, Data Entry, Excel, PDF, Data Analysis, Data Management · I have between one and five PDF files that contain purely numerical data laid out in complex, nested table structures. I need every figure lifted out of those PDFs and placed into an Excel workbook so the spreadsheet mirrors the original hierarchy—parent rows, indented subtotals, and any multi-level groupings must stay intact.

Accuracy is everything: each value should land in the correct row and column, with no dropped decimals or shifted totals. If you prefer using automated tools such as Python-pdfplumber, Camelot, Tabula, or Adobe Acrobat batch scripts, that’s fine, but manual checks are mandatory to guarantee the final sheet reconciles perfectly with the source.

Deliverables • One Excel file per PDF (or a single multi-sheet workbook if that’s cleaner) reproducing the full table layout. • A brief note outlining the extraction method you used and any assumptions made.

I will spot-check against the PDFs, so please keep your interim work available until I confirm everything balances.