FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Tech Stack
Tools & technologiesAWSCloudDockerJavaScriptPythonSelenium
About the role
Key responsibilities & impact- Handle specialized data scraping tasks for the Tendem project
- Own end-to-end data extraction workflows across complex websites
- Ensure complete coverage, accuracy, and reliable delivery of structured datasets
- Use tools and custom workflows to accelerate data collection, validation, and task execution
- Extract data reliably from dynamic and interactive web sources
- Adapt scraping approaches for JavaScript-rendered content and changing site behavior
- Enforce data quality through validation checks, cross-source consistency controls, formatting specifications, and systematic verification
- Scale scraping operations for large datasets using efficient batching or parallelization
- Monitor failures and maintain stability against minor site structure changes
Requirements
What you’ll need- At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development
- Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus
- Strong technical foundation and practical experience with scripting, automation, and data extraction workflows
- Experience solving non-trivial problems and using modern development tools and technologies
- Experience systematically collecting, structuring, and validating data from diverse sources
- Strong experience in Python web scraping, including BeautifulSoup, Selenium or similar, dynamic content (JS, AJAX, infinite scroll), and APIs via proxies
- Ability to extract data from complex structures, including hierarchies, archived pages, and inconsistent HTML
- Experience in data cleaning, normalization, and validation, delivering structured datasets in CSV, JSON, or Google Sheets
- Experience handling anti-bot mechanisms and dynamic site structures at scale
- Experience with cloud infrastructure such as AWS or equivalent and containerization with Docker
- Hands-on experience with LLM frameworks such as LangChain, OpenRouter, or similar applied to automation tasks
- Strong attention to detail and commitment to data accuracy
- Self-directed work ethic and ability to troubleshoot independently
- English proficiency at Upper-intermediate (B2) or above
- A GitHub link is a plus
Benefits
Comp & perks- Freelance opportunity
- Part-time remote work
- Estimated workload of around 10–20 hours per week during active project phases
