Jobs · OTHR

Freelance Data Scraping Engineer (Python)

Mindrift · Iowa, United States · 4 wk ago
RemoteRemoteOTHR$37/hrPart-time

About the role

The Mindrift platform connects specialists with innovative technology projects. Our mission is to help develop high-quality AI technologies by combining real-world expertise from professionals across the globe with advanced AI development efforts.

Responsibilities

  • Own end-to-end data extraction workflows across complex websites, ensuring complete coverage, accuracy, and reliable delivery of structured datasets
  • Leverage available tools and custom workflows to accelerate data collection, validation, and task execution while meeting defined requirements
  • Ensure reliable extraction from dynamic and interactive web sources, adapting approaches as needed to handle JavaScript-rendered content and changing site behavior
  • Enforce data quality standards through validation checks, cross-source consistency controls, adherence to formatting specifications, and systematic verification prior to delivery
  • Scale scraping operations for large datasets using efficient batching or parallelization, monitor failures, and maintain stability against minor site structure changes

Requirements

  • At least 3+ years of relevant experience in data engineering, web scraping, automation, or software development
  • Strong experience in Python web scraping (BeautifulSoup, Selenium or similar), including dynamic content (JS, AJAX, infinite scroll) and APIs via proxies
  • Proven ability to extract data from complex structures (hierarchies, archived pages, inconsistent HTML)
  • Solid background in data cleaning, normalization, and validation, delivering structured datasets (CSV, JSON, Google Sheets)

Qualifications

  • Bachelor's or Master's Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields (plus)
  • Hands-on experience with LLMs and AI frameworks to enhance automation and problem-solving
  • Strong attention to detail and commitment to data accuracy
  • Self-directed work ethic with ability to troubleshoot independently

Technical Skills

  • Strong experience in Python web scraping (BeautifulSoup, Selenium or similar), including dynamic content (JS, AJAX, infinite scroll) and APIs via proxies
  • Proven ability to extract data from complex structures (hierarchies, archived pages, inconsistent HTML)
  • Solid background in data cleaning, normalization, and validation, delivering structured datasets (CSV, JSON, Google Sheets)

Additional Requirements

  • A link to GitHub is a plus

Pay

Earn up to $37 per hour equivalent, depending on your level and pace of contribution.

Schedule

Tasks are estimated to require around 10-20 hours per week during active phases, based on project requirements. This is an estimate, not a guaranteed workload, and applies only while the project is active.

Similar jobs