FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in end-to-end data extraction workflows, ensuring data quality and reliability while utilizing advanced web scraping techniques and tools. Proficient in handling dynamic web content and large datasets, with a strong focus on automation and data validation.
Highest-signal resume keywords
Python Web ScrapingBeautifulSoupSeleniumData Cleaning and ValidationAWS or Equivalent Cloud Infrastructure
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data ExtractionWeb ScrapingAutomationData NormalizationDynamic Content HandlingAPIs via ProxiesData Structuring in CSV, JSON, Google SheetsAnti-Bot MechanismsContainerization with DockerLLM Frameworks
Soft Skills
Attention to DetailSelf-Directed Work EthicTroubleshooting
Tools & Technologies
ApifyOpenRouterGitHub
Industry Keywords
Data EngineeringWeb ScrapingSoftware DevelopmentTechnical Fields
Tech Stack
Tools & technologiesAWSCloudDockerJavaScriptPythonSelenium
About the role
Key responsibilities & impact- Own end-to-end data extraction workflows across complex websites
- Ensure complete coverage, accuracy, and reliable delivery of structured datasets
- Use available tools and custom workflows to accelerate data collection, validation, and task execution
- Extract data reliably from dynamic and interactive web sources, adapting to JavaScript-rendered content and changing site behavior
- Enforce data quality through validation checks, cross-source consistency controls, formatting specifications, and systematic verification
- Scale scraping operations for large datasets using efficient batching or parallelization
- Monitor failures and maintain stability against minor site structure changes
- Apply tools such as Apify, OpenRouter, and other technologies alongside technical expertise
Requirements
What you’ll need- At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development
- Strong experience in Python web scraping, including BeautifulSoup, Selenium or similar, dynamic content (JS, AJAX, infinite scroll), and APIs via proxies
- Proven ability to extract data from complex structures such as hierarchies, archived pages, and inconsistent HTML
- Solid background in data cleaning, normalization, and validation, delivering structured datasets in CSV, JSON, or Google Sheets
- Demonstrated experience handling anti-bot mechanisms and dynamic site structures at scale
- Experience with cloud infrastructure such as AWS or equivalent and containerization with Docker
- Hands-on experience with LLM frameworks such as LangChain, OpenRouter, or similar, applied to automation tasks
- Strong attention to detail and commitment to data accuracy
- Self-directed work ethic with ability to troubleshoot independently
- English proficiency at Upper-intermediate (B2) or above
- Bachelor's or Master's Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus
- GitHub link is a plus
Benefits
Comp & perks- Freelance opportunity
- Part-time workload estimated at 10–20 hours per week during active project phases
- Remote work
- Access to innovative technology projects through the Mindrift platform
