FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in data extraction workflows, web scraping, and data processing, ensuring high-quality, structured datasets while adapting to dynamic web environments. Proficient in Python and familiar with cloud infrastructure and containerization for scalable data operations.
Highest-signal resume keywords
Python Web ScrapingData ExtractionData ValidationCloud Infrastructure (AWS)Containerization (Docker)
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Web ScrapingData ProcessingData CleaningData NormalizationDynamic Content HandlingAPIs via ProxiesAnti-Bot MechanismsHierarchical Data ExtractionCSV FormattingJSON Formatting
Soft Skills
Attention to DetailSelf-Directed Work EthicTroubleshooting
Tools & Technologies
BeautifulSoupSeleniumApifyOpenRouterLangChainDocker
Industry Keywords
Data EngineeringAutomationSoftware DevelopmentDynamic Web SourcesStructured Datasets
Tech Stack
Tools & technologiesAWSCloudDockerJavaScriptPythonSelenium
About the role
Key responsibilities & impact- Own end-to-end data extraction workflows across complex websites, ensuring complete coverage, accuracy, and reliable delivery of structured datasets
- Leverage available tools and custom workflows to accelerate data collection, validation, and task execution while meeting defined requirements
- Ensure reliable extraction from dynamic and interactive web sources, adapting approaches for JavaScript-rendered content and changing site behavior
- Enforce data quality standards through validation checks, cross-source consistency controls, formatting specifications, and systematic verification before delivery
- Scale scraping operations for large datasets using efficient batching or parallelization
- Monitor failures and maintain stability against minor site structure changes
- Collaborate with Tendem Agents handling repetitive tasks within Mindrift’s hybrid AI + human system
- Use tools such as Apify, OpenRouter, and other technologies alongside technical expertise and custom approaches
- Apply web scraping, data extraction, and data processing expertise to deliver accurate, reliable, high-quality results
Requirements
What you’ll need- At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development (required)
- Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus
- Strong experience in Python web scraping (BeautifulSoup, Selenium or similar), including dynamic content (JS, AJAX, infinite scroll) and APIs via proxies
- Proven ability to extract data from complex structures, including hierarchies, archived pages, and inconsistent HTML
- Solid background in data cleaning, normalization, and validation, delivering structured datasets in CSV, JSON, or Google Sheets
- Demonstrated experience handling anti-bot mechanisms and dynamic site structures at scale
- Experience with cloud infrastructure (AWS or equivalent) and containerization (Docker) as part of real workflows
- Hands-on experience with LLM frameworks (LangChain, OpenRouter, or similar) applied to automation tasks
- Strong attention to detail and commitment to data accuracy
- Self-directed work ethic with ability to troubleshoot independently
- English proficiency: Upper-intermediate (B2) or above (required)
- A link to GitHub is a plus
Benefits
Comp & perks- Freelance opportunity
- Part-time remote work
- Estimated 10–20 hours per week during active project phases
