About this job
<p><strong>Mindrift is looking for highly skilled Senior Python Data Scraping Engineers to join the Tendem project and drive specialized data scraping workflows within our hybrid AI + human system. </strong></p>
<p>In this role, as an AI Pilot – that’s how we refer to this role at Mindrift – you’ll collaborate with Tendem Agents that handle repetitive tasks, while you provide critical thinking, domain expertise, and quality control to deliver accurate and actionable results. </p>
<p>This part-time remote opportunity is ideal for technical professionals with hands-on experience in web scraping, data extraction and processing.</p>
<p><strong>What We Do</strong></p>
<p>The Mindrift platform connects specialists with AI projects from major tech innovators. Our mission is to unlock the potential of Generative AI by tapping into real-world expertise from across the globe.</p>
<p>This is a freelance role for a Tendem project. As a <strong>Senior Python Data Scraping Engineer</strong>, you'll handle data scraping tasks requiring technical precision for web extraction and processing, utilizing various tools such as our provided Apify and OpenRouter alongside your own resourceful approaches. </p>
<p><strong>Key Responsibilities:</strong></p>
<ul>
<li>Own end-to-end data extraction workflows across complex websites, ensuring complete coverage, accuracy, and reliable delivery of structured datasets.</li>
<li>Leverage internal tools (Apify, OpenRouter) alongside custom workflows to accelerate data collection, validation, and task execution while meeting defined requirements.</li>
<li>Ensure reliable extraction from dynamic and interactive web sources, adapting approaches as needed to handle JavaScript-rendered content and changing site behavior.</li>
<li>Enforce data quality standards through validation checks, cross-source consistency controls, adherence to formatting specifications, and systematic verification prior to delivery.</li>
<li>Scale scraping operations for large datasets using efficient batching or parallelization, monitor failures, and maintain stability against minor site structure changes.</li>
</ul>
<p><strong>Requirements:</strong></p>
<ul>
<li>At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development (required).</li>
<li>Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus.</li>
<li>Candidates should have a strong technical foundation and practical experience with scripting, automation, and AI-assisted workflows. We are looking for specialists who can solve non-trivial problems, work confidently with LLMs, and systematically collect, structure, and validate data from diverse sources. A methodical, detail-oriented approach and the ability to work independently are essential.</li>
<li>Strong experience in Python web scraping (BeautifulSoup, Selenium or similar), including dynamic content (JS, AJAX, infinite scroll) and APIs via proxies</li>
<li>Proven ability to extract data from complex structures (hierarchies, archived pages, inconsistent HTML)</li>
<li>Solid background in data cleaning, normalization, and validation, delivering structured datasets (CSV, JSON, Google Sheets)</li>
<li>Demonstrated experience handling anti-bot mechanisms and dynamic site structures at scale</li>
<li>Experience with cloud infrastructure (AWS or equivalent) and containerization (Docker) as part of real workflows</li>
<li>Hands-on experience with LLM frameworks (LangChain, OpenRouter, or similar) applied to automation tasks</li>
<li>Strong attention to detail and commitment to data accuracy</li>
<li>Self-directed work ethic with ability to troubleshoot independently</li>
<li>A link to GitHub is a plus</li>
<li>English proficiency: Upper-intermediate (B2) or above (required)</li>
</ul>
<p><strong>Project time expectations</strong></p>
<p>For this project, tasks are estimated to require around 10–20 hours per week during active phases, based on project requirements. This is an estimate, not a guaranteed workload, and applies only while the project is active.</p>
<p><strong>Compensation </strong></p>
<p>On this project, contributors can earn up to <strong>$45 per hour equivalent</strong>, depending on their level and pace of contribution. </p>
<p>Compensation varies across projects depending on scope, complexity, and required expertise. Please note that other projects on the platform may offer different earning levels based on their requirements.</p>