Spectrum Health

Data Engineer

Spectrum Health$95K — $105K *
Healthcare
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in Computer Science, Software Engineering, Data Analytics, or a related field.
  • 3-7 years of practical data engineering experience, especially in data migration and cloud architectures.
  • Proficient in ER diagramming and various database modeling techniques (3NF for OLTP, Star/Snowflake for OLAP).
  • Experience with ETL platforms like Informatica and Azure Data Factory for data extraction and transformation.
  • Familiarity with Azure Data Services including ADLS Gen2, Synapse, and Databricks, plus streaming technologies such as Event Hubs and Kafka.
  • Knowledge of Salesforce data structures and experience using SOQL and Salesforce Bulk API.
  • Advanced SQL proficiency for database queries, schema design, and performance optimization.

Responsibilities

  • Design, deploy, and manage a centralized Azure ADLS Gen2 or Snowflake data lake to act as the single source of truth.
  • Create both batch and real-time ETL/ELT processes for ingesting EHR data and operational telemetry using technologies like Kafka and Spark Streaming.
  • Develop and execute complex migration scripts with tools like Informatica and Azure Data Factory to transition data from various sources into the lakehouse.
  • Implement automated processes for validation, deduplication, and encryption to ensure data integrity and compliance with PHI regulations.
  • Build high-performance data pipelines that synchronize data with the Salesforce Health Cloud using various integration tools.
  • Transform healthcare data standards (HL7, FHIR, EDI) into organized formats suitable for storage and analysis.
  • Design comprehensive data models for various analytics needs, including OLTP and OLAP systems.

Benefits

  • Opportunities for professional development and continuous learning.
  • Engagement in high-impact projects, contributing to healthcare data management innovations.
  • Collaboration within a dynamic team focused on data-driven healthcare solutions.
Full Job Description
Job Description
  • Design, deploy and maintain a centralized Azure ADLS Gen2/Databricks/Synapse (or Snowflake) data lake that serves as the single source of truth
  • Create batch & real-time ETL/ELT flows (Azure Event Hubs, Kafka, Spark Streaming) to ingest EHR data and operational telemetry
  • Write and run complex migration scripts using Informatica PowerCenter/IDMC or Azure Data Factory to move data from Epic, other EHRs and legacy OLTP stores into the lakehouse and Salesforce Health Cloud
  • Build automated validation, deduplication, masking, encryption and column-level security to guarantee 100% integrity and PHI compliance
  • Build high-throughput pipelines (SOQL, Bulk API 2.0, Informatica, MuleSoft, dbt, Python) that sync Health Cloud data with the central repository
  • Transform HL7, FHIR and EDI messages into clean canonical schemas for storage and analytics
  • Design Robust Data Models - Perform ER/3NF modeling for OLTP and dimensional (star/snowflake/medallion) modeling for OLAP analytics


Qualifications
  • Bachelor's degree in Computer Science, Software Engineering, Data Analytics, or a related technical field
  • Experience: 3-7 years of hands-on data engineering experience, with a proven track record of executing complex data migration projects and building cloud data lake architectures.
  • Expertise in ER (Entity-Relationship) diagramming, relational database modeling (3NF for OLTP), and dimensional modeling (Star Schema, Snowflake, Medallion) for OLAP systems
  • Experience with enterprise ETL platforms such as Informatica (Informatica PowerCenter / IDMC), Azure Data Factory, SSIS, or Talend for legacy data extraction and transformation
  • Experience with cloud data architectures (Azure data services: Azure Data Factory, ADLS Gen2, Azure Synapse, Databricks) and familiarity with streaming/real-time pipelines (Event Hubs, Kafka, Spark Streaming)
  • Experience working with Salesforce data structures, SOQL, Salesforce Bulk API, and data loading tools (e.g.,Informatica Salesforce Data Loader, MuleSoft, Fivetran, dbt, or custom Python scripts)
  • Expert-level SQL skills (complex joins, CTEs, window functions, schema design, index tuning, and database optimization)
  • Strong Python skills for data manipulation, scripting, and pipeline execution using libraries like Pandas, PySpark, and SQLAlchemy
  • Compensation is based on experience (Annual salary range $95,000 - $105, 000)

#corp_IT

Additional Information

This position is a current vacancy

About Spectrum Health

Spectrum Health is a not-for-profit, integrated health system based in West Michigan. Our organization includes a medical center, regional community hospitals, a dedicated children's hospital, a multispecialty medical group and a nationally recognized health plan, Priority Health. We invest in our people, technologies and facilities to create a high-quality, sustainable, service-oriented and cost-effective system of health care.
Learn more about Spectrum Health
Size
31,000 employees
Industry

Similar Jobs

More Jobs at Spectrum Health

  • Spectrum Health
    Data Engineer
    $95K — $105K *
    Toronto, ON M3C 0E3
    Healthcare
    In-Person

More Healthcare Jobs

Find similar Data Engineer jobs: