Creative Information Technology

Data Engineer - Baltimore City, MD

Creative Information Technology$110K — $130K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's or master's degree in Computer Science, Statistics, Mathematics, Economics, or related field; equivalent experience accepted.
  • 3+ years of hands-on experience building and maintaining data pipelines on AWS or similar platforms.
  • Strong proficiency in Python and SQL; Scala or Java knowledge is a plus.
  • Proven experience with Apache Spark (PySpark) for large-scale data processing.
  • Familiarity with AWS services such as Glue, S3, Redshift, Athena, EMR, and Lake Formation.
  • Experience with EDI X12, HL7, or FHIR data formats.
  • Understanding of HIPAA and CMS compliance requirements.

Responsibilities

  • Design and maintain data pipelines and ETL processes for structured and unstructured data.
  • Build data storage solutions including data lakehouses and data warehouses to support analytics.
  • Develop data reliability and quality processes.
  • Prepare data for modeling purposes and monitor data processing systems.
  • Collaborate with cross-functional teams to gather requirements and objectives.
  • Troubleshoot performance and reliability issues within data systems.
  • Create and maintain documentation for built processes.

Benefits

  • Opportunity to work with cutting-edge technologies in cloud data engineering.
  • Collaborative work environment with cross-functional teams.
  • Focus on healthcare data integration providing impactful work.
  • Possibility to contribute to CI/CD and automation initiatives.
  • Support for continuous learning and adopting new technologies.
Full Job Description
Data Engineer - Baltimore City, MD

Background

Client is seeking a hands-on Data Engineer to design, develop, and optimize large-scale data pipelines in support of our Enterprise Data Warehouse (EDW) and Data Lake solutions. This role requires deep technical expertise in coding, pipeline orchestration, and cloud-native data engineering on AWS. The Data Engineer will be directly responsible for implementing ingestion, transformation, and integration workflows - ensuring data is high-quality, compliant, and analytics-ready. This role may support other projects or teams within MDH as needed.

Role and Responsibilities

Responsible for designing, building, and maintaining data pipelines and infrastructure to support data-driven decisions and analytics. The individual is responsible for the following tasks:

  • Design, develop and maintain data pipelines, and extract, transform, load (ETL) processes to collect, process and store structured and unstructured data
  • Build data architecture and storage solutions, including data lakehouses, data lakes, data warehouse, and data marts to support analytics and reporting
  • Develop data reliability, efficiency, and qualify checks and processes
  • Prepare data for data modeling
  • Monitor and optimize data architecture and data processing systems
  • Collaboration with multiple teams to understand requirements and objectives
  • Administer testing and troubleshooting related to performance, reliability, and scalability

  • Create and update documentation


Hands-On Data Pipeline Development

  • Design, code, and deploy ETL/ELT pipelines across bronze, silver, and gold layers of the Data Lakehouse.
  • Build ingestion pipelines for structured (SQL), semi-structured (JSON, XML), and unstructured data using PySpark/Python programming language using AWS Glue or EMR.
  • Implement incremental loads, deduplication, error handling, and data validation.
  • Actively troubleshoot, debug, and optimize pipelines for scalability and cost efficiency.


EDW & Data Lake Implementation

  • Develop dimensional data models (Star Schema, Snowflake Schema) for analytics and reporting.
  • Build and maintain tables in Iceberg, Delta Lake, or equivalent OTF formats.
  • Optimize partitioning, indexing, and metadata for fast query performance.


Healthcare Data Integration

  • Build ingestion and transformation pipelines for EDI X12 transactions (837, 835, 278, etc.).
  • Implement mapping and transformation of EDI data with FHIR and HL7 frameworks.
  • Work hands-on with AWS Health Lake (or equivalent) to store and query healthcare data.


Data Quality, Security & Compliance

  • Develop automated validation scripts to enforce data quality and integrity.
  • Implement IAM roles, encryption, and auditing to meet HIPAA and CMS compliance standards.
  • Maintain lineage and governance documentation for all pipelines.


Collaboration & Delivery

  • Work closely with the Lead Data Engineer, analysts, and data scientists to deliver pipelines that support enterprise-wide analytics.
  • Actively contribute to CI/CD pipelines, Infrastructure-as-Code (IaC), and automation.
  • Continuously improve pipelines and adopt new technologies where appropriate.


Minimum Qualifications

  • The candidate should have experience as data engineer or similar role with a strong understanding of data architecture and ETL processes. The candidate should be proficient in programming languages for data processing and knowledgeable of distributed computing and parallel processing.
  • This position requires a bachelor's or master's degree from an accredited college or university with a major in computer science, statistics, mathematics, economics, or a related field. Three (3) years of equivalent experience in a related field may be substituted for the Bachelor's degree.
  • 3+ years hands-on experience in building, deploying, and maintaining data pipelines on AWS or equivalent cloud platforms.
  • Strong coding skills in Python and SQL (Scala or Java a plus).
  • Proven experience with Apache Spark (PySpark) for large-scale processing.
  • Hands-on experience with AWS Glue, S3, Redshift, Athena, EMR, Lake Formation.
  • Strong debugging and performance optimization skills in distributed systems.
  • Hands-on experience with Iceberg, Delta Lake, or other OTF table formats.
  • Experience with Airflow or other pipeline orchestration frameworks.
  • Practical experience in CI/CD and Infrastructure-as-Code (Terraform, CloudFormation).
  • Practical experience with EDI X12, HL7, or FHIR data formats.
  • Strong understanding of Medallion Architecture for data lake houses.
  • Hands-on experience building dimensional models and data warehouses.
  • Working knowledge of HIPAA and CMS interoperability requirements.

About Creative Information Technology

Creative Information Technology is a software development company that specializes in custom software development, web development, and mobile app development. The company was founded in 1998 and is headquartered in Wilmington, Delaware. Creative Information Technology's clients include small and medium-sized businesses across a range of industries.
Learn more about Creative Information Technology
Size
100 employees
Industry
Founded
1998

Similar Jobs

  • Rochester Regional Health Systems
    EMR Architect - Data
    $100K — $130K *
    Rochester Regional Health Systems
    Remote
  • Data Engineer
    $125K — $155K *
    Alpha Opco LLC
    White Plains, NY 10605 (Westchester County)
  • Full Stack Data Engineer
    $95K — $115K *
    Cavaliers Holdings LLC
    Cleveland, OH 44130 (Cuyahoga County)
  • Data Engineer
    $77K — $176K *
    Booz Allen Hamilton, Inc.
    Aberdeen Proving Ground, MD 21005 (Harford County)
  • Data Engineer
    $77K — $176K *
    Booz Allen Hamilton, Inc.
    Belcamp, MD 21017 (Harford County)
  • Belay Technologies
    Junior Software Engineer
    $70K — $190K *
    Belay Technologies
    Columbia, MD 21044 (Howard County)

More Jobs at Creative Information Technology

More Information Technology Jobs

Find similar Data Engineer - Baltimore City, MD jobs: