Cigna-Evernorth Services Inc. seeks Software Engineering Advisors for the Plano, TX location to build and maintain scalable data pipelines in Databricks to deliver high-quality, timely data for client deliverables.
Responsibilities:
• Design and implement modular ETL workflows using PySpark for data ingestion, cleansing, transformation and optimizing large volumes of structured medical data.
• Monitor production jobs and provide operational support to troubleshoot data failures, ensuring system reliability and business continuity.
• Participate in the delivery of enterprise-level data solutions that support strategic decision-making via analytics and reporting.
• Serve as subject matter expert for critical information Management assignments, collaborating with stakeholders across departments.
• Design and develop consolidated conformed enterprise data warehouse and data lake to store critical data across Customer, Provider, Claims, Client, and Benefits data.
• Analyze data to extract actionable insights, improve data usability, and ensure alignment with client-specific reporting and business requirements.
• Optimize data models and transformation logic to enhance data accuracy and relevance for downstream analytics.
• Apply statistical techniques and domain knowledge to interpret healthcare accurately, ensuring its relevance and usability for business teams and leadership.
• Develop and manage CI/CD pipelines using tools like GitHub Actions to automate code integration, testing, deployment and release management for data workflows.
• Ensure data compliance with healthcare regulations, including HIPAA and PHI handling guidelines, during all stages of data processing.
• Translate business requirements into technical solutions by understanding and visualizing complex data flows.
• Aligning outcomes with client goals.
• Hybrid work schedule.
Qualifications:
• Master's degree in Data Science, Computer Science or closely related field and 3 years of experience designing ETL pipelines. Will accept a Bachelor's degree and 7 years of experience.
• Must have experience with:
• Designing and developing robust ETL pipelines using PySpark, Scala, SQL, and Shell scripting
• PostgreSQL, MySQL, or SQL Server
• Cassandra
• Snowflake
• Using multiple file formats in data pipelines, including Parquet, CSV, and ORC
• Big Data technologies such as Hadoop, Hive, Spark, or Kafka
• AWS tools such as Glue, IAM, EC2, EMR, RDS, S3, Athena, Step Functions, or Lambda
• Data mapping, validation, and modeling
• Building and maintaining CI/CD pipelines using GitHub or GitLab
• Reporting and BI tools such as Tableau or Power BI
• Working with Delta tables in Databricks
• RabbitMQ for queuing high-speed sensor data
• Natural Language Processing (NLP) techniques
• DevOps tools such as Terraform, Git, and Jira
If you will be working at home occasionally or permanently, the internet connection must be obtained through a cable broadband or fiber optic internet service provider with speeds of at least 10Mbps download/5Mbps upload.