About the Role
This is a freelance role for the Tendem project, focused on specialized data scraping workflows for real-world use cases. You will apply expertise in web scraping, data extraction, and data processing to deliver accurate, reliable, and high-quality results using tools like Apify and OpenRouter.
Responsibilities
- Own end-to-end data extraction workflows across complex websites, ensuring complete coverage, accuracy, and reliable delivery of structured datasets.
- Leverage available tools and custom workflows to accelerate data collection, validation, and task execution while meeting defined requirements.
- Ensure reliable extraction from dynamic and interactive web sources, adapting approaches as needed to handle JavaScript-rendered content and changing site behavior.
- Enforce data quality standards through validation checks, cross-source consistency controls, adherence to formatting specifications, and systematic verification prior to delivery.
- Scale scraping operations for large datasets using efficient batching or parallelization, monitor failures, and maintain stability against minor site structure changes.
Requirements
- At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development.
- Candidates should have a strong technical foundation and practical experience with scripting, automation, and data extraction workflows.
- Specialists who can solve non-trivial problems, work confidently with modern development tools and technologies, and systematically collect, structure, and validate data from diverse sources.
- A methodical, detail-oriented approach and the ability to work independently are essential.
- Demonstrated experience handling anti-bot mechanisms and dynamic site structures at scale.
- Experience with cloud infrastructure (AWS or equivalent) and containerization (Docker) as part of real workflows.
- Hands-on experience with LLM frameworks (LangChain, OpenRouter, or similar) applied to automation tasks.
- Strong attention to detail and commitment to data accuracy.
- Self-directed work ethic with ability to troubleshoot independently.
- English proficiency: Upper-intermediate (B2) or above.
Skills
- Python web scraping (BeautifulSoup, Selenium or similar)
- Dynamic content scraping (JS, AJAX, infinite scroll)
- API data extraction via proxies
- Data extraction from complex structures (hierarchies, archived pages, inconsistent HTML)
- Data cleaning, normalization, and validation
- Structured dataset delivery (CSV, JSON, Google Sheets)
- LLM frameworks (LangChain, OpenRouter, or similar) applied to automation tasks
Location
- Remote
Work Type
- Part-time
- Freelance
Experience Level
- Senior
- 5+ years of relevant experience
Education Level
- Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus.
Salary/Compensations
- Up to $45 per hour equivalent
About the Company
- The Mindrift platform connects specialists with innovative technology projects.
- Our mission is to help develop high-quality AI technologies by combining real-world expertise from professionals across the globe with advanced AI development efforts.
