Senior Python Data Scraping Engineer – Freelance

Posted 21hrs ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Senior Python engineer building scalable web-scraping workflows for Mindrift’s AI technology projects. Extracting, validating, and processing structured data from complex dynamic websites.

Responsibilities:

  • Join the Tendem project and drive specialized data scraping workflows for real-world applications
  • Own end-to-end data extraction workflows across complex websites
  • Ensure complete coverage, accuracy, and reliable delivery of structured datasets
  • Use available tools and custom workflows to accelerate data collection, validation, and task execution
  • Extract data reliably from dynamic and interactive web sources, adapting to JavaScript-rendered content and changing site behavior
  • Enforce data quality through validation checks, cross-source consistency controls, formatting specifications, and systematic verification
  • Scale scraping operations for large datasets using efficient batching or parallelization
  • Monitor failures and maintain stability against minor site structure changes
  • Utilize tools including Apify, OpenRouter, and other technologies alongside technical expertise and custom approaches

Requirements:

  • At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development (required)
  • Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus
  • Strong technical foundation and practical experience with scripting, automation, and data extraction workflows
  • Strong experience in Python web scraping, including BeautifulSoup, Selenium or similar, dynamic content (JS, AJAX, infinite scroll), and APIs via proxies
  • Proven ability to extract data from complex structures, including hierarchies, archived pages, and inconsistent HTML
  • Solid background in data cleaning, normalization, and validation, delivering structured datasets in CSV, JSON, or Google Sheets
  • Demonstrated experience handling anti-bot mechanisms and dynamic site structures at scale
  • Experience with cloud infrastructure (AWS or equivalent) and containerization (Docker) as part of real workflows
  • Hands-on experience with LLM frameworks such as LangChain, OpenRouter, or similar applied to automation tasks
  • Strong attention to detail and commitment to data accuracy
  • Self-directed work ethic with ability to troubleshoot independently
  • A link to GitHub is a plus
  • English proficiency: Upper-intermediate (B2) or above (required)

Benefits:

  • Freelance opportunity
  • Part-time remote work
  • Estimated 10–20 hours per week during active project phases
  • Access to innovative technology projects through the Mindrift platform