FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Data Engineer II – School of Medicine
Emory University. Develop and evolve the Unified Data Platform (UDP), built on Azure Fabric .
Tech Stack
Tools & technologiesAirflowApacheAzureETLPySparkPythonSparkSQL
About the role
Key responsibilities & impact- Develop and evolve the Unified Data Platform (UDP), built on Azure Fabric
- Establish collaborative relationships with subject matter experts and understand the line of business
- Collaborate with Business Analysts, Project Managers, Data Analysts, Architects, and other cross-functional teams
- Work with researchers and stakeholders to gather, analyze, and translate requirements into technical solutions
- Design, develop, and maintain scalable data pipelines, ETL/ELT processes, and data integration workflows
- Build and optimize solutions integrating disparate data sources while ensuring data quality, integrity, and consistency
- Evaluate emerging technologies and develop proof-of-concepts
- Apply biomedical informatics standards, methodologies, and principles to research data solutions
- Ensure adherence to HIPAA and institutional data governance policies and standards
- Develop and maintain metadata, data standards, and data governance processes for complex datasets
- Create and maintain technical documentation for data pipelines, workflows, and system designs
- Communicate technical concepts to technical and non-technical stakeholders
- Partner with data and system architects and apply data modeling best practices
- Manage workload and report task progress and deliverables in a timely manner
- Develop complex reports, data pipelines, and ETL processes from disparate systems and ensure accuracy
- Perform other related duties as required
Requirements
What you’ll need- Applicants must be legally authorized to work in the United States
- Position is not eligible for visa sponsorship now or in the future
- Bachelor's degree in a related field and three years of related experience, OR an equivalent combination of education, training, and experience
- Strong experience designing and maintaining data pipelines and ETL/ELT processes
- Proficiency in Python, SQL, and Apache Spark (PySpark)
- Experience with modern data platforms such as Azure Fabric, Azure Synapse Serverless, or Databricks
- Ability to work with large, complex, and distributed datasets
- Solid understanding of data modeling concepts and best practices
- Familiarity with data governance, metadata management, and data quality frameworks
- Knowledge of HIPAA and healthcare data compliance standards
- Experience with CI/CD practices and version control systems such as Git
- Familiarity with data pipeline orchestration and monitoring tools such as Airflow or Azure Data Factory
- Understanding of biomedical informatics methodologies is a plus
- Exposure to FHIR, OMOP, or similar healthcare data models is nice-to-have
- Strong analytical, problem-solving, and critical-thinking abilities
- Excellent communication and collaboration skills across diverse teams
- Ability to adapt quickly to new technologies and evolving data environments
Benefits
Comp & perks- Remote work classification: Full Remote – Monthly
- Tasks can be performed remotely with only occasional visits to an Emory University location
- Equal opportunity employer
- Reasonable accommodations for qualified individuals with disabilities upon request