FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in end-to-end data extraction workflows, ensuring data quality and accuracy through validation and systematic verification. Proficient in Python web scraping and cloud infrastructure, with a strong focus on handling dynamic content and anti-bot mechanisms.
Highest-signal resume keywords
Python Web ScrapingData Cleaning and ValidationCloud Infrastructure (AWS)Dynamic Content ExtractionAutomation with LLM Frameworks
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data Extraction WorkflowsWeb ScrapingBeautifulSoupSeleniumAPIs via ProxiesData NormalizationCSV and JSON FormattingAnti-Bot MechanismsContainerization with DockerData Structuring
Soft Skills
Attention to DetailSelf-Directed Work EthicTroubleshooting Ability
Tools & Technologies
AWSDockerLangChainOpenRouterGoogle Sheets
Industry Keywords
Data EngineeringWeb ScrapingAutomationSoftware DevelopmentDynamic Site Structures
Tech Stack
Tools & technologiesAWSCloudDockerJavaScriptPythonSelenium
About the role
Key responsibilities & impact- Own end-to-end data extraction workflows across complex websites
- Ensure complete coverage, accuracy, and reliable delivery of structured datasets
- Use tools and custom workflows to accelerate data collection, validation, and task execution
- Extract data from dynamic and interactive web sources, including JavaScript-rendered content
- Adapt scraping approaches to changing site behavior and structures
- Enforce data quality through validation checks, cross-source consistency controls, formatting compliance, and systematic verification
- Scale scraping operations for large datasets using efficient batching or parallelization
- Monitor failures and maintain stability against minor site structure changes
- Work on specialized data scraping workflows for real-world applications in the Tendem project
Requirements
What you’ll need- At least 5+ years of relevant experience in data engineering, web scraping, automation, or software development (required)
- Bachelor’s or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus
- Strong experience in Python web scraping, including BeautifulSoup, Selenium or similar, dynamic content, and APIs via proxies
- Proven ability to extract data from complex structures such as hierarchies, archived pages, and inconsistent HTML
- Solid background in data cleaning, normalization, and validation, delivering structured datasets in CSV, JSON, or Google Sheets
- Demonstrated experience handling anti-bot mechanisms and dynamic site structures at scale
- Experience with cloud infrastructure such as AWS or equivalent and containerization with Docker
- Hands-on experience with LLM frameworks such as LangChain, OpenRouter, or similar applied to automation tasks
- Strong attention to detail and commitment to data accuracy
- Self-directed work ethic and ability to troubleshoot independently
- English proficiency: Upper-intermediate (B2) or above (required)
- GitHub link is a plus
Benefits
Comp & perks- Freelance opportunity
- Part-time remote work
- Estimated 10–20 hours per week during active phases
- Up to $30 per hour equivalent, depending on level and pace of contribution
- Access to innovative technology projects through the Mindrift platform
