Extract PDF Data to Excel

Employer not named by the sourceRemote

Full Stack

Apply on the company’s site

Frontier is not the employer and does not collect applications.

About this role

Python, Data Processing, Data Entry, Excel, OCR, Data Extraction, Data Analysis, Data Management · A batch of multi-page PDF files needs to become a clean, analysis-ready Excel workbook. Each text element—whether it appears as straight digital text or inside a scanned image—has to land in the correct cell so the spreadsheet mirrors the logical flow of the original documents.

The job is straightforward but accuracy is critical: no dropped characters, no shifted columns, no merged cells where they don’t belong. I’m happy for you to use any reliable method—Python (tabula-py, camelot, pdfplumber), Power Query, Acrobat automation, ABBYY FineReader OCR, or a manual approach—so long as the end result is an .xlsx file I can filter and pivot without cleanup.

Deliverables • One Excel workbook containing every record from the supplied PDFs, organised consistently and ready for immediate use. • A brief note (or script) describing the extraction method so the process is reproducible if new PDFs arrive later.

Final file must be double-checked for completeness; spot checks should prove 100 % coverage and less than 1 % transcription error. If that sounds routine to you, let’s get started.