Senior Data Acquisition & Document Intelligence Engineer
Employer not named by the sourceRemote
Full StackSenior engineering jobs
Frontier is not the employer and does not collect applications.
About this role
Python, Data Processing, Web Scraping, OCR, Scrapy, Data Scraping, Data Extraction, Data Management, Data Engineer, OCR Automation · Project Title: Senior Data Acquisition & Document Intelligence Engineer
We are looking for a highly experienced Senior Data Acquisition & Document Intelligence Engineer to build, operate, and continuously improve a large-scale production data acquisition and document processing system.
Project Type: Long-term / Ongoing Engagement: Full-time Experience Required: 5+ years Location: On-site / Hybrid
About the Project
We collect large volumes of publicly available documents and structured records from hundreds of external web sources every day.
The sources are highly inconsistent and frequently change their website structure, formats, URLs, APIs, and document layouts. A significant portion of the data also comes from scanned PDFs, images, and other difficult-to-process documents.
We need someone who can take end-to-end ownership of the entire data acquisition pipeline, including web scraping, crawling, document downloading, OCR, data extraction, structuring, validation, quality assurance, monitoring, and ongoing maintenance.
This is not a one-time scraping project. We are looking for someone with proven experience building and maintaining production-grade extraction pipelines at