Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Avanade

Senior Data Engineer

Avanade

. Design, build, and maintain batch and/or streaming data pipelines using Python/PySpark and SQL .

Posted 9/17/2026full-timeSão Paulo • BrazilSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Proficient in designing and maintaining data pipelines using Python, PySpark, and SQL, with a strong focus on ETL/ELT processes and data modeling for both OLTP and OLAP systems. Demonstrates expertise in software architecture best practices, version control with GitHub, and implementing robust testing strategies.

Highest-signal resume keywords
Python ProficiencyPySpark ExpertiseSQL KnowledgeETL/ELT Process DevelopmentGitHub Version Control

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Data Pipeline DesignData ModelingETL/ELT ProcessesUnit TestingIntegration TestingData Quality TestingSoftware Architecture Best PracticesQuery TuningData TransformationDistributed Processing
Soft Skills
CollaborationDocumentationContinuous Improvement
Tools & Technologies
AirflowDatabricksAzureGitHub ActionsData Lake StorageSynapseFabricNoSQL DatabasesGraph DatabasesColumnar Databases
Certifications & Qualifications
DP-203AZ-900DP-900Databricks Lakehouse FundamentalsData Engineer Associate Certification
Industry Keywords
OLTPOLAPData GovernanceRBACABACData QualitySchema ValidationPerformance OptimizationAutomationStandardization

Tech Stack

Tools & technologies
AirflowAzureETLNoSQLPySparkPythonSQLVault

About the role

Key responsibilities & impact
  • Design, build, and maintain batch and/or streaming data pipelines using Python/PySpark and SQL
  • Develop and orchestrate robust, observable ETL/ELT processes with logging, metrics, and alerts
  • Model data across transactional (OLTP) and analytical/multidimensional (OLAP) layers for BI, Analytics, and AI
  • Ensure code and artifact quality, versioning, and reproducibility using GitHub, including branches, pull requests, and code reviews
  • Implement unit, integration, and data tests, and plan end-to-end testing strategies
  • Apply software architecture best practices, including modularity, security, performance, and cost efficiency
  • Collaborate with analysts, data scientists, and product teams to translate business requirements into scalable solutions
  • Document data, processes, and architectural decisions to promote knowledge sharing
  • Support continuous improvement through automation, standardization, and performance optimization of data environments

Requirements

What you’ll need
  • Proficiency in Python and PySpark for data manipulation, transformation, and distributed processing
  • Strong knowledge of SQL, including data modeling, query tuning, CTEs, window functions, and partitioning
  • Experience with transactional (OLTP) and multidimensional (OLAP) data modeling
  • Experience developing ETL/ELT processes and orchestration using tools such as Airflow, Databricks Jobs, or similar
  • Knowledge of software architecture applied to data, including design patterns, security, scalability, and cost efficiency
  • Experience with GitHub, including GitFlow, pull requests, semantic versioning, and basic GitHub Actions
  • Experience developing and planning tests, including data tests such as data quality and schema validation
  • Preferred: experience with CI/CD pipelines for data, such as GitHub Actions or Azure DevOps, automated testing, and job/notebook deployment
  • Preferred: experience with non-relational data stores and architectures, such as data lakes, NoSQL, document, columnar, and graph databases
  • Preferred: experience with cloud platforms, preferably Azure, including Data Lake Storage, Databricks, Synapse, Fabric, and Key Vault
  • Preferred: exposure to data governance and security, including lineage, catalogs, and RBAC/ABAC
  • A plus: Microsoft Azure certifications such as DP-203, AZ-900, DP-900, or equivalent
  • A plus: Databricks Lakehouse Fundamentals or Data Engineer Associate/Professional certifications
  • Intermediate to advanced English proficiency

Benefits

Comp & perks
  • Meal or food allowance
  • Flexible benefits card (up to Senior Consultant level)
  • Medical and dental insurance
  • Certifications and training
  • Life insurance
  • Private pension plan
  • Avababy: pregnancy support and a welcome kit for new parents
  • Company profit-sharing plan
  • Wellhub
  • Childcare assistance
  • Career mentoring
  • Birthday Off policy for you and your children up to 12 years old
  • Wellness sessions
  • For management positions: company vehicle, parking, and fuel allowance