FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in data extraction workflows, ensuring data quality and accuracy while utilizing advanced tools and technologies for web scraping. Proficient in Python and familiar with cloud infrastructure and containerization to handle complex data challenges.
Highest-signal resume keywords
Python Web ScrapingData ValidationCloud Infrastructure (AWS)Data Extraction WorkflowsAutomation Tools (Selenium, BeautifulSoup)
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data ExtractionData CleaningData StructuringScriptingAutomationDynamic Content HandlingAPI IntegrationData NormalizationBatch ProcessingParallelization
Soft Skills
Attention to DetailSelf-Directed Work EthicProblem Solving
Tools & Technologies
ApifyOpenRouterDockerGitHubLangChain
Industry Keywords
Web ScrapingData EngineeringAutomationSoftware DevelopmentData Quality Assurance
Tech Stack
Tools & technologiesAWSCloudDockerJavaScriptPythonSelenium
About the role
Key responsibilities & impact- Handle end-to-end data extraction workflows across complex websites
- Ensure complete coverage, accuracy, and reliable delivery of structured datasets
- Use available tools and custom workflows to accelerate data collection, validation, and task execution
- Extract data reliably from dynamic and interactive web sources, adapting to JavaScript-rendered content and changing site behavior
- Enforce data quality through validation checks, cross-source consistency controls, formatting specifications, and systematic verification
- Scale scraping operations for large datasets using efficient batching or parallelization
- Monitor failures and maintain stability against minor site structure changes
- Use tools such as Apify, OpenRouter, and other technologies alongside technical expertise and custom approaches
- Work on specialized data scraping workflows for real-world applications as part of the Tendem project
Requirements
What you’ll need- At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development
- Strong technical foundation and practical experience with scripting, automation, and data extraction workflows
- Experience solving non-trivial problems and using modern development tools and technologies
- Experience systematically collecting, structuring, and validating data from diverse sources
- Strong experience in Python web scraping, including BeautifulSoup, Selenium or similar, dynamic content (JS, AJAX, infinite scroll), and APIs via proxies
- Ability to extract data from complex structures, including hierarchies, archived pages, and inconsistent HTML
- Background in data cleaning, normalization, and validation, delivering structured datasets in CSV, JSON, or Google Sheets
- Experience handling anti-bot mechanisms and dynamic site structures at scale
- Experience with cloud infrastructure such as AWS or equivalent and containerization with Docker
- Hands-on experience with LLM frameworks such as LangChain, OpenRouter, or similar applied to automation tasks
- Strong attention to detail and commitment to data accuracy
- Self-directed work ethic and ability to troubleshoot independently
- English proficiency at Upper-intermediate (B2) or above
- GitHub link is a plus
- Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus
Benefits
Comp & perks- Part-time freelance opportunity
- Remote work
- Estimated 10–20 hours per week during active project phases
- Contributors can earn up to $45 per hour equivalent, depending on level and pace of contribution
