Senior Data Acquisition & Document Intelligence Engineer

Employer not named by the sourceRemote

Full StackSenior engineering jobs

Apply on the company’s site

Frontier is not the employer and does not collect applications.

About this role

Python, Data Processing, Web Scraping, OCR, Scrapy, Data Scraping, Data Extraction, Data Management, Data Engineer, OCR Automation · Project Title: Senior Data Acquisition & Document Intelligence Engineer

We are looking for a highly experienced Senior Data Acquisition & Document Intelligence Engineer to build, operate, and continuously improve a large-scale production data acquisition and document processing system.

Project Type: Long-term / Ongoing Engagement: Full-time Experience Required: 5+ years Location: On-site / Hybrid

About the Project

We collect large volumes of publicly available documents and structured records from hundreds of external web sources every day.

The sources are highly inconsistent and frequently change their website structure, formats, URLs, APIs, and document layouts. A significant portion of the data also comes from scanned PDFs, images, and other difficult-to-process documents.

We need someone who can take end-to-end ownership of the entire data acquisition pipeline, including web scraping, crawling, document downloading, OCR, data extraction, structuring, validation, quality assurance, monitoring, and ongoing maintenance.

This is not a one-time scraping project. We are looking for someone with proven experience building and maintaining production-grade extraction pipelines at