FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and maintaining data lake pipelines using Apache Spark, Scala, and SQL, while ensuring data quality and performance optimization. Proficient in workflow orchestration with tools like Control-M and Airflow, and experienced in collaborating with cross-functional teams to deliver high-quality data solutions.
Highest-signal resume keywords
Apache SparkScalaSQLControl-MAirflow
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data Lake ArchitectureData Processing SolutionsData IngestionData TransformationData ValidationPerformance OptimizationRoot-Cause AnalysisCI/CD PracticesGitHubJenkins
Soft Skills
Analytical SkillsProblem-Solving SkillsCommunication SkillsTeamwork SkillsAdaptability
Tools & Technologies
DatabricksAzureAWSYARNServiceNow
Industry Keywords
Market RiskP&LValuationsRisk MetricsRegulatory Reporting
Tech Stack
Tools & technologiesAirflowApacheAWSAzureCloudJenkinsScalaServiceNowSparkSQLYarn
About the role
Key responsibilities & impact- Design, develop and maintain data lake pipelines using Apache Spark, Scala and SQL
- Build and maintain scalable data processing solutions for large volumes of data
- Orchestrate batch workflows and dependencies using tools such as Control-M and Airflow
- Implement data ingestion, transformation, validation and publication processes for Market Risk use cases
- Analyse and optimise the performance, reliability and scalability of Spark-based workloads
- Analyse technical and functional requirements, identify gaps and propose effective solutions
- Contribute to delivery planning and collaborate with different teams to ensure successful implementation
- Support production incidents, perform root-cause analysis and contribute to continuous improvement of operational processes
- Work closely with Agile squads, functional teams and production support teams to deliver high-quality solutions and business value
- Follow software engineering and data development best practices to ensure robust, maintainable and production-ready solutions
Requirements
What you’ll need- 2+ years of professional experience in Big Data, Data Lakes or cloud-based data platforms
- Strong hands-on experience with Apache Spark, Scala and SQL for large-scale data processing
- Good understanding of Data Lake architecture, data quality controls, data lineage and production-grade data pipelines
- Strong analytical and problem-solving skills
- Excellent communication and teamwork skills, with the ability to collaborate effectively with technical and functional stakeholders
- Adaptability and a proactive approach to solving technical challenges
- English: C1 level or equivalent professional fluency
- Spanish is considered a plus
- Experience with workflow orchestration and scheduling tools, particularly Control-M and Airflow, highly valued
- Experience with Databricks and cloud technologies such as Azure and AWS highly valued
- Experience with YARN and cloud-based compute environments highly valued
- Experience with CI/CD and software engineering practices, including GitHub, Jenkins and SonarQube, highly valued
- Experience with banking production environments, operational monitoring and incident management tools such as ServiceNow highly valued
- Knowledge of Market Risk processes, including P&L, valuations, risk metrics and regulatory reporting data flows highly valued
Benefits
Comp & perks- 100% Remote set up
- Training and career development
- Possibility to be part of a multicultural team and work on international projects
