FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and developing scalable data pipelines and infrastructure, with a strong focus on data ingestion, transformation, and storage solutions. Proven ability to lead technical teams, mentor engineers, and collaborate effectively across departments to deliver high-quality data solutions.
Highest-signal resume keywords
Python ProgrammingApache SparkApache AirflowCloud-Based Data InfrastructureData Modeling
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data Pipeline DevelopmentDistributed Data ProcessingDatabase DesignQuery OptimizationData Feature Lifecycle ManagementData RecoveryMonitoring and ObservabilityVersion ControlData Lake ArchitectureSoftware Design Best Practices
Soft Skills
CommunicationMentoringProblem SolvingCollaborationPersuasion
Tools & Technologies
AWSGCPDatabricksData Science ToolsAnalytics Platforms
Industry Keywords
Data EngineeringHigh-Performance Data SolutionsProduction DatasetsTechnical LeadershipData Infrastructure
Tech Stack
Tools & technologiesAirflowApacheAWSCloudGoogle Cloud PlatformPySparkPythonSpark
About the role
Key responsibilities & impact- Lead the design and development of scalable, high-performance data pipelines and infrastructure powering Samba TV’s analytics
- Resolve complex technical issues across disciplines
- Lead the design, build, and maintenance of high-scale production datasets
- Ensure delivery of versioned outputs and reliable customer-facing reports
- Manage data feature lifecycles, including attributes, business logic, safe rollouts, backfills, and reprocessing
- Architect efficient data ingestion, transformation, and storage solutions
- Improve performance, reliability, security, and compute and storage costs
- Serve as technical lead for production incidents, investigate root causes, implement fixes, and validate data recovery
- Advise contacts and colleagues on complex matters
- Mentor engineers and guide development of policies and ideas
- Enhance monitoring and observability of data processes
- Collaborate with Data Science, Analytics, and Product teams on production-ready data solutions
Requirements
What you’ll need- Typically 8+ years of related experience with a Bachelor’s degree, or 6 years with a Master’s, or 3 years with a PhD
- Advanced knowledge of Python
- Deep understanding of distributed data processing frameworks like Apache Spark or PySpark
- Expertise in Apache Airflow and Databricks
- Extensive experience with cloud-based data infrastructure (AWS, GCP) and modern data lake architectures
- Strong knowledge of data modeling, database design, and query optimization for relational and non-relational databases
- Proven track record of best practices for code quality, testing, and software design in a data engineering context
- Ability to adapt communication style and use persuasion when delivering messages related to the wider firm business
Benefits
Comp & perks- Bonuses
- Short-term incentives
- Long-term incentives
- Health insurance
- Wellness offerings
- Life and disability insurance
- Retirement savings plan
- Paid holidays
- Paid time off (PTO)
