FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Tech Stack
Tools & technologiesAWSCloudDockerJavaScriptPythonSelenium
About the role
Key responsibilities & impact- Handle end-to-end data extraction workflows across complex websites
- Ensure complete coverage, accuracy, and reliable delivery of structured datasets
- Use tools and custom workflows to accelerate data collection, validation, and task execution
- Extract data from dynamic and interactive web sources, including JavaScript-rendered content
- Adapt scraping approaches to changing site behavior
- Enforce data quality standards through validation checks and cross-source consistency controls
- Follow formatting specifications and systematically verify data before delivery
- Scale scraping operations for large datasets using efficient batching or parallelization
- Monitor failures and maintain stability against minor site structure changes
- Apply tools such as Apify, OpenRouter, and other technologies alongside technical expertise
Requirements
What you’ll need- At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development
- Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus
- Strong technical foundation and practical experience with scripting, automation, and data extraction workflows
- Experience solving non-trivial problems and working with modern development tools and technologies
- Ability to systematically collect, structure, and validate data from diverse sources
- Methodical, detail-oriented approach
- Ability to work independently
- Strong experience in Python web scraping using BeautifulSoup, Selenium or similar
- Experience handling dynamic content including JS, AJAX, and infinite scroll
- Experience working with APIs via proxies
- Ability to extract data from complex structures such as hierarchies, archived pages, and inconsistent HTML
- Experience in data cleaning, normalization, and validation
- Ability to deliver structured datasets in CSV, JSON, and Google Sheets
- Demonstrated experience handling anti-bot mechanisms and dynamic site structures at scale
- Experience with cloud infrastructure such as AWS or equivalent
- Experience with containerization using Docker
- Hands-on experience with LLM frameworks such as LangChain, OpenRouter, or similar applied to automation tasks
- Strong attention to detail and commitment to data accuracy
- Self-directed work ethic and ability to troubleshoot independently
- English proficiency at Upper-intermediate (B2) or above
- A GitHub link is a plus
Benefits
Comp & perks- Freelance opportunity
- Part-time remote work
- Estimated 10–20 hours per week during active project phases
- Compensation up to $25 per hour equivalent, depending on level and pace of contribution
