Extract Formatted PDF Text to Excel
Employer not named by the sourceRemote
Frontier is not the employer and does not collect applications.
About this role
Python, Data Processing, Data Entry, Excel, PDF, ABBYY FineReader, Adobe Acrobat, Data Management · I have a collection of PDFs that contain only flowing paragraphs of text. I need every word lifted into a single Excel workbook while preserving the original feel—paragraph breaks, line indents, bold, italics, and any other inline styling that appears in the source. There are no tables, forms, or images involved, so the job is entirely text-focused.
Use whatever mix of tools you prefer—Adobe Acrobat, ABBYY FineReader, python-pdfplumber with openpyxl, or straightforward manual copy-paste—provided the finished file opens cleanly in Microsoft Excel and shows the same formatting cues I see in the PDFs.
Please ensure: • Each PDF’s content is clearly separated (one worksheet or a plainly labeled section). • No characters are lost, scrambled, or replaced. • All visible formatting from the PDFs is retained, not approximated.
Deliverable: one .xlsx file containing the fully formatted text from every PDF.
Let me know how quickly you can turn this around and what method you plan to use so I can green-light the workflow right away.