Extract PDF Tables to Excel

Employer not named by the sourceRemote

Full Stack

Apply on the company’s site

Frontier is not the employer and does not collect applications.

About this role

Python, Data Processing, Software Architecture, PDF, Data Analysis, Data Management · I have a PDF in which every page contains tables laid out in exactly the same way. I need those tables lifted out and delivered in an Excel workbook organised over multiple worksheets—one per logical section rather than everything crammed into a single sheet.

Accuracy of every cell matters because the file will feed straight into an existing reporting model. Please preserve column order, numeric formats, and any header rows that appear in the source. If you prefer scripting the task, Python with tabula-py or camelot is fine, and I’m equally happy with a manual approach if the result is 100 % correct.

Once complete, send back: • The clean, fully-formatted .xlsx file with the data split across the required sheets • A brief note on the method you used so I can reproduce the extraction if the PDF gets updated in future.

Let me know your turnaround time and any clarifying questions you might have.