About the Role
Join the Tendem project as a Senior Python Data Scraping Engineer to drive specialized data scraping workflows within a hybrid AI + human system. As an AI Pilot, collaborate with Tendem Agents, providing critical thinking, domain expertise, and quality control for accurate, actionable results. This part-time remote role is for technical professionals experienced in web scraping, data extraction, and processing.
Responsibilities
- Own end-to-end data extraction workflows across complex websites, ensuring complete coverage, accuracy, and reliable delivery of structured datasets.
- Leverage internal tools (Apify, OpenRouter) alongside custom workflows to accelerate data collection, validation, and task execution while meeting defined requirements.
- Ensure reliable extraction from dynamic and interactive web sources, adapting approaches as needed to handle JavaScript-rendered content and changing site behavior.
- Enforce data quality standards through validation checks, cross-source consistency controls, adherence to formatting specifications, and systematic verification prior to delivery.
- Scale scraping operations for large datasets using efficient batching or parallelization, monitor failures, and maintain stability against minor site structure changes.
Requirements
- At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development.
- Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus.
- Strong technical foundation and practical experience with scripting, automation, and AI-assisted workflows.
- Ability to solve non-trivial problems.
- Confidence working with LLMs.
- Ability to systematically collect, structure, and validate data from diverse sources.
- Methodical, detail-oriented approach.
- Ability to work independently.
- Strong experience in Python web scraping (BeautifulSoup, Selenium or similar), including dynamic content (JS, AJAX, infinite scroll) and APIs via proxies.
- Proven ability to extract data from complex structures (hierarchies, archived pages, inconsistent HTML).
- Solid background in data cleaning, normalization, and validation, delivering structured datasets (CSV, JSON, Google Sheets).
- Demonstrated experience handling anti-bot mechanisms and dynamic site structures at scale.
- Experience with cloud infrastructure (AWS or equivalent) and containerization (Docker) as part of real workflows.
- Hands-on experience with LLM frameworks (LangChain, OpenRouter, or similar) applied to automation tasks.
- Strong attention to detail and commitment to data accuracy.
- Self-directed work ethic with ability to troubleshoot independently.
- A link to GitHub is a plus.
- English proficiency: Upper-intermediate (B2) or above.
Skills
- Python web scraping
- BeautifulSoup
- Selenium
- Data extraction from complex structures
- Data cleaning
- Data normalization
- Data validation
- Handling anti-bot mechanisms
- Cloud infrastructure (AWS)
- Containerization (Docker)
- LLM frameworks (LangChain)
- LLM frameworks (OpenRouter)
- Scripting
- Automation
- AI-assisted workflows
- Critical thinking
- Domain expertise
- Quality control
- Problem-solving
- Working with LLMs
- Data collection
- Data structuring
- Attention to detail
- Independent troubleshooting
- English proficiency (Upper-intermediate B2)
Location
- Remote
Work Type
- Part-time
- Remote
- Freelance
Experience Level
- Senior
- 5+ years of relevant experience
Education Level
- Bachelor's Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields (plus)
- Master's Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields (plus)
Salary/Compensations
- Up to $45 per hour equivalent
About the Company
- Mindrift connects specialists with AI projects from major tech innovators. Its mission is to unlock Generative AI potential by leveraging global real-world expertise.
