Senior Python Data Scraping Engineer – Freelance

Posted 20hrs ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Senior Python engineer building scalable web-scraping workflows for Mindrift’s AI technology projects. Extracting, validating, and processing structured data from complex, dynamic websites.

Responsibilities:

  • Own end-to-end data extraction workflows across complex websites
  • Ensure complete coverage, accuracy, and reliable delivery of structured datasets
  • Leverage available tools and custom workflows to accelerate data collection, validation, and task execution
  • Ensure reliable extraction from dynamic and interactive web sources
  • Adapt approaches to JavaScript-rendered content and changing site behavior
  • Enforce data quality through validation checks, cross-source consistency controls, formatting specifications, and systematic verification
  • Scale scraping operations for large datasets using efficient batching or parallelization
  • Monitor failures and maintain stability against minor site structure changes
  • Use tools such as Apify, OpenRouter, and other technologies for web extraction and processing

Requirements:

  • At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development
  • Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus
  • Strong technical foundation and practical experience with scripting, automation, and data extraction workflows
  • Strong experience in Python web scraping, including BeautifulSoup, Selenium or similar, dynamic content (JS, AJAX, infinite scroll), and APIs via proxies
  • Proven ability to extract data from complex structures, including hierarchies, archived pages, and inconsistent HTML
  • Solid background in data cleaning, normalization, and validation, delivering structured datasets in CSV, JSON, and Google Sheets
  • Demonstrated experience handling anti-bot mechanisms and dynamic site structures at scale
  • Experience with cloud infrastructure (AWS or equivalent) and containerization (Docker) as part of real workflows
  • Hands-on experience with LLM frameworks such as LangChain, OpenRouter, or similar applied to automation tasks
  • Strong attention to detail and commitment to data accuracy
  • Self-directed work ethic with ability to troubleshoot independently
  • English proficiency: Upper-intermediate (B2) or above
  • A link to GitHub is a plus

Benefits:

  • Freelance project opportunity
  • Part-time schedule estimated at 10–20 hours per week during active project phases
  • Remote work
  • Compensation up to $40 per hour equivalent, depending on level and pace of contribution