EPAM Systems

Senior Data Software Engineer/ Databricks, Apache Spark, PySpark

EPAM Systems$120K — $145K *
US-AnywhereRemote in Georgia, US
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 3+ years of experience in Data Software Engineering
  • Proficient in Apache Spark, preferably with Databricks
  • Strong skills in PySpark and SQL
  • Experience with unit testing frameworks, specifically pytest
  • Familiar with data engineering principles and ETL processes
  • Knowledge of version control systems like Git
  • Upper-Intermediate (B2) or higher English proficiency

Responsibilities

  • Design and develop scalable data pipelines using Apache Spark
  • Write optimized PySpark code for data transformation
  • Execute complex SQL queries for extraction and validation
  • Implement unit tests using pytest for code quality
  • Collaborate with team members for high-quality data solutions
  • Monitor and troubleshoot data workflows and performance
  • Document technical designs and best practices

Benefits

  • Opportunity to work with leading-edge technology in data engineering
  • Collaborative team environment with data scientists and analysts
  • Emphasis on code quality with a strong unit testing culture
  • Potential exposure to cloud platforms and DevOps practices
  • Professional development through hands-on projects and documentation practices
Full Job Description
We are seeking a skilled Senior Data Software Engineer with strong expertise in PySpark , SQL , and unit testing to join our data engineering team. The ideal candidate will have hands-on experience with Apache Spark , preferably within the Databricks environment, and will be responsible for building scalable data pipelines, optimizing data workflows, and ensuring code quality through rigorous testing practices. Responsibilities Design, develop, and maintain scalable data pipelines using Apache Spark (Databricks preferred) Write efficient and optimized PySpark code for data transformation and processing Develop and execute complex SQL queries for data extraction, validation, and reporting Implement unit tests using pytest to ensure code reliability and maintainability Collaborate with data scientists, analysts, and other engineers to deliver high-quality data solutions Monitor and troubleshoot data workflows and performance issues Document technical designs, processes, and best practices Requirements 3+ years of experience in Data Software Engineering Proven experience with Apache Spark, ideally in a Databricks environment Proficiency in PySpark and SQL Background in unit testing frameworks, especially pytest Understanding of data engineering principles and ETL processes Familiarity with version control systems (e.g., Git) Ability to work independently and in a collaborative team setting Excellent problem-solving and communication skills Proficiency in English at an Upper-Intermediate level (B2) or higher Nice to have Experience with cloud platforms (e.g., Azure, AWS, GCP) Knowledge of CI/CD pipelines and DevOps practices Familiarity with Delta Lake, MLflow, or other Databricks-native tools

About EPAM Systems

EPAM Systems, Inc. is a leading global provider of digital platform engineering and development services. The company has a strong presence in North America, Europe, and Asia, and serves clients in a variety of industries, including financial services, healthcare, and retail. EPAM's services include software engineering, product development, and digital platform engineering, and the company has a reputation for delivering high-quality solutions that help its clients achieve their business goals. EPAM has been recognized as a leader in the digital services industry by a number of independent research firms, and the company has won numerous awards for its work.
Learn more about EPAM Systems
Size
58,824 employees
Market Cap
$18.2 billion
Industry
Net Income
$327.1 million
Founded
1993
5 Year Trend
+26.5%
Revenue
$2.6 billion
NASDAQ

Similar Jobs

More Jobs at EPAM Systems

More Information Technology Jobs

Find similar Senior Data Software Engineer/ Databricks, Apache Spark, PySpark jobs: