FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and maintaining data lake pipelines using Apache Spark, Scala, and SQL, while ensuring data quality and performance optimization. Proficient in collaborating with cross-functional teams and applying best practices in software engineering and data development.
Highest-signal resume keywords
Apache SparkScalaSQLControl-MAirflow
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data Lake ArchitectureData Processing SolutionsData IngestionData TransformationData ValidationData PublicationPerformance OptimizationRoot-Cause AnalysisCI/CD PracticesAgile Methodologies
Soft Skills
Analytical SkillsProblem-Solving SkillsCommunication SkillsTeamwork SkillsAdaptability
Tools & Technologies
DatabricksAzureAWSYARNGitHubJenkinsSonarQubeServiceNow
Industry Keywords
Big DataMarket RiskP&LValuationsRisk MetricsRegulatory Reporting
Tech Stack
Tools & technologiesAirflowApacheAWSAzureCloudJenkinsScalaServiceNowSparkSQLYarn
About the role
Key responsibilities & impact- Design, develop and maintain data lake pipelines using Apache Spark, Scala and SQL
- Build and maintain scalable data processing solutions for large volumes of data
- Orchestrate batch workflows and dependencies using Control-M and Airflow
- Implement data ingestion, transformation, validation and publication processes for Market Risk use cases
- Analyse and optimise Spark workload performance, reliability and scalability
- Analyse technical and functional requirements, identify gaps and propose effective solutions
- Contribute to delivery planning and collaborate with different teams
- Support production incidents and perform root-cause analysis
- Contribute to continuous improvement of operational processes
- Collaborate with Agile squads, functional teams and production support teams
- Apply software engineering and data development best practices to deliver robust, maintainable and production-ready solutions
Requirements
What you’ll need- 2+ years of professional experience in Big Data, Data Lakes or cloud-based data platforms
- Strong hands-on experience with Apache Spark, Scala and SQL for large-scale data processing
- Good understanding of Data Lake architecture, data quality controls, data lineage and production-grade data pipelines
- Strong analytical and problem-solving skills
- Excellent communication and teamwork skills
- Ability to collaborate with technical and functional stakeholders
- Adaptability and proactive approach to solving technical challenges
- English: C1 level or equivalent professional fluency
- Experience with workflow orchestration and scheduling tools, particularly Control-M and Airflow, highly valued
- Experience with Databricks, Azure and AWS highly valued
- Experience with YARN and cloud-based compute environments highly valued
- Experience with CI/CD and software engineering practices, including GitHub, Jenkins and SonarQube, highly valued
- Experience with banking production environments, operational monitoring and incident management tools such as ServiceNow highly valued
- Knowledge of Market Risk processes, including P&L, valuations, risk metrics and regulatory reporting data flows, highly valued
Benefits
Comp & perks- Full-time, permanent employment contract
- 100% remote working with the flexibility to work from wherever you are
- Professional growth and career development opportunities
- Access to continuous learning
- Multicultural and international environment
- Private medical insurance
- Life insurance
- Lunch and transport cards through the flexible remuneration programme
- Online language classes
- Smart Office Pack to support remote working
