Emory University

Data Engineer II (School of Medicine)

Emory University • $95K — $115K *
Healthcare
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in a related field or equivalent experience
  • 3+ years of related data engineering experience
  • Strong experience with data pipelines and ETL/ELT processes
  • Proficiency in Python, SQL, and PySpark
  • Familiarity with Azure Fabric, Azure Synapse, or Databricks
  • Understanding of data governance and HIPAA compliance
  • Excellent communication skills and ability to collaborate across teams

Responsibilities

  • Establish collaborative relationships with subject matter experts
  • Collaborate with cross-functional teams to deliver solutions
  • Gather and analyze requirements from researchers and stakeholders
  • Design and maintain scalable data pipelines and ETL processes
  • Ensure data quality and integrity across multiple systems
  • Evaluate emerging technologies and create proof-of-concepts
  • Apply biomedical informatics methodologies to research solutions

Benefits

  • Remote work capability with occasional site visits to Emory University
  • Work within business hours of the Eastern time zone
  • Dynamic work environment with evolving technologies
  • Collaborative team culture focused on innovation
  • Opportunity to develop skills in advanced data engineering tools
Full Job Description
Description

The Data Engineer II is a key contributor to the development and evolution of the Unified Data Platform (UDP), built on Azure Fabric. This role works closely with researchers, architects, and cross-functional teams to design and deliver scalable, high-quality data solutions that support research, analytics, and clinical insights. The position involves working with complex datasets, modern cloud data platforms, and emerging technologies. While experience with Azure Fabric is preferred, candidates with strong hands-on experience in Azure Synapse or Databricks with PySpark will be equally considered.

This role requires a strong foundation in data engineering, biomedical informatics principles, data governance, and compliance standards such as HIPAA. The ideal candidate is a collaborative team player who can translate business and research requirements into efficient and scalable technical solutions.

*Applicants must be legally authorized to work in the United States. This position is not eligible for visa sponsorship now or in the future.

JOB DESCRIPTION:
  • Establishes collaborative relationships with subject matter experts and develops an understanding of the line of business.
  • Collaborate as a core member of cross-functional teams including Business Analysts, Project Managers, Data Analysts, and Architects to deliver high-quality solutions within scope and timeline.
  • Work directly with researchers and stakeholders to gather, analyze, and translate requirements into technical solutions.
  • Design, develop, and maintain scalable data pipelines, ETL/ELT processes, and data integration workflows across multiple systems.
  • Build and optimize data solutions that integrate data from disparate sources, ensuring data quality, integrity, and consistency.
  • Evaluate emerging technologies and develop proof-of-concepts to support innovation and continuous improvement.
  • Apply biomedical informatics standards, methodologies, and principles to research data solutions.
  • Ensure adherence to HIPAA and institutional data governance policies and standards.
  • Develop and maintain metadata, data standards, and data governance processes for complex datasets.
  • Create and maintain clear technical documentation, including data pipelines, workflows, and system designs.
  • Communicate technical concepts effectively to both technical and non-technical stakeholders.
  • Partner with data and system architects and demonstrate an understanding of data modeling concepts and best practices.
  • Manage workload effectively and provide timely updates on task progress and deliverables.
  • Works as a positive team member of a project that may consist of Business Analysts, Project Managers, Information Architects, Data Analysts, and/or Database Administrators to deliver quality applications and components within scope, on time, and within budget.
  • Manages workload effectively and report status of tasks in a timely manner.
  • Works directly with researchers to document, analyze, and translate their needs into technical designs and informatics solutions.
  • Participates in the evaluation of emerging technologies and develops proof-of-concepts.
  • Contributes to technical teams.
  • Follows standard operational procedures and HIPAA regulations.
  • Develops strategies for managing complex data sets through maintaining data standards and metadata.
  • Applies biomedical informatics technical standards, methodologies, and principles to research-specific program needs, objectives, and outcomes.
  • Develops complex reports, data pipelines, and ETL processes from disparate systems and ensures their accuracy.
  • Gathers user requirements and creates technical documentation
  • Performs other related duties as required.

MINIMUM QUALIFICATIONS:
  • A bachelor's degree in a related field and three years of related experience, OR an equivalent combination of education, training, and experience.

PREFERRED QUALIFICATIONS:
  • Strong experience designing and maintaining data pipelines and ETL/ELT processes
  • Proficiency in Python, SQL, and Apache Spark (PySpark)
  • Experience with modern data platforms such as Azure Fabric (preferred), Azure Synapse Serverless, or Databricks
  • Ability to work with large, complex, and distributed datasets
  • Solid understanding of data modeling concepts and best practices
  • Familiarity with data governance, metadata management, and data quality frameworks
  • Knowledge of HIPAA and healthcare data compliance standards
  • Experience with CI/CD practices and proficiency with version control systems like Git
  • Familiarity with data pipeline orchestration and monitoring tools Like Airflow , Azure Data Factory or similar frameworks
  • Understanding of biomedical informatics methodologies is a plus
  • Strong analytical, problem-solving, and critical-thinking abilities
  • Excellent communication and collaboration skills across diverse teams
  • Ability to adapt quickly to new technologies and evolving data environments
  • Exposure to FHIR, OMOP, or similar healthcare data models (nice-to-have)

NOTE: Tasks related to this position can be performed remotely with only occasional visits to an Emory University location. Eastern (EST) time zone business hours may apply. Emory reserves the right to change this status with notice to employee. Emory does not approve as a primary work location in the following states; NJ, AK, and HI, any U.S. Territories or outside of the United States.

About Emory University

Emory University is a private research university in Atlanta, Georgia. The university was founded in 1836 as Emory College and has since grown into a leading research institution with nine academic divisions. Emory University is known for its liberal arts college, professional schools, and biomedical research. The university has a diverse student body and offers undergraduate, graduate, and professional degree programs. Emory University is consistently ranked among the top 25 universities in the United States and is a member of the Association of American Universities.
Learn more about Emory University
Size
30,000 employees
Industry

Similar Jobs

More Jobs at Emory University

More Healthcare Jobs

Find similar Data Engineer II (School of Medicine) jobs: