FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in end-to-end data extraction workflows, ensuring data accuracy and reliability while utilizing advanced web scraping techniques and tools. Proficient in handling complex data structures and maintaining data quality through systematic validation and cleaning processes.
Highest-signal resume keywords
Python Web ScrapingBeautifulSoupSeleniumData Cleaning and ValidationAWS Cloud Infrastructure
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data ExtractionWeb ScrapingAutomationData NormalizationDynamic Content HandlingAPIs via ProxiesData Structuring in CSVJSONGoogle SheetsContainerization with Docker
Soft Skills
Attention to DetailSelf-Directed Work EthicTroubleshooting
Tools & Technologies
ApifyOpenRouterLangChainGitHub
Industry Keywords
Data EngineeringWeb ScrapingAutomationDynamic Site StructuresAnti-Bot Mechanisms
Tech Stack
Tools & technologiesAWSCloudDockerJavaScriptPythonSelenium
About the role
Key responsibilities & impact- Handle end-to-end data extraction workflows across complex websites
- Ensure complete coverage, accuracy, and reliable delivery of structured datasets
- Use tools such as Apify, OpenRouter, and custom workflows to accelerate data collection, validation, and task execution
- Extract data reliably from dynamic and interactive web sources, including JavaScript-rendered content
- Adapt scraping approaches to changing site behavior and minor site structure changes
- Enforce data quality through validation checks, cross-source consistency controls, formatting specifications, and systematic verification
- Scale scraping operations for large datasets using efficient batching or parallelization
- Monitor failures and maintain scraping stability
- Work on specialized data scraping workflows for the Tendem project at Mindrift
Requirements
What you’ll need- At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development
- Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus
- Strong experience in Python web scraping, including BeautifulSoup, Selenium or similar, dynamic content, and APIs via proxies
- Proven ability to extract data from complex structures, including hierarchies, archived pages, and inconsistent HTML
- Solid background in data cleaning, normalization, and validation, delivering structured datasets in CSV, JSON, and Google Sheets
- Demonstrated experience handling anti-bot mechanisms and dynamic site structures at scale
- Experience with cloud infrastructure such as AWS or equivalent
- Experience with containerization using Docker
- Hands-on experience with LLM frameworks such as LangChain, OpenRouter, or similar, applied to automation tasks
- Strong attention to detail and commitment to data accuracy
- Self-directed work ethic and ability to troubleshoot independently
- English proficiency: Upper-intermediate (B2) or above
- GitHub link is a plus
Benefits
Comp & perks- Freelance opportunity
- Part-time remote work
- Estimated 10–20 hours per week during active project phases
- Access to innovative technology projects through the Mindrift platform
