EPAM Systems

Lead Data Software Engineer with Databricks

EPAM Systems$130K — $155K *
US-AnywhereRemote in Georgia, US
Enterprise Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years of experience in data analytics or platform roles with leadership responsibilities
  • Expertise in cloud-based data platforms, particularly Azure
  • Hands-on experience with Databricks, Spark/PySpark, and SQL
  • Knowledge of Python and data manipulation libraries such as Pandas and NumPy
  • Familiarity with data visualization tools like Power BI, Tableau, and Looker
  • Experience in data quality frameworks and governance
  • Understanding of data governance, security, and access control

Responsibilities

  • Lead the design and management of traceability data platforms and pipelines
  • Define and implement architectural standards and best practices
  • Ensure data quality and integrity through robust monitoring frameworks
  • Support data engineers with troubleshooting and optimization
  • Drive continuous improvement of data quality and auditability
  • Translate regulatory requirements into actionable technical solutions
  • Collaborate with cross-functional teams for alignment and scalability

Benefits

  • Opportunity to drive technical innovation in traceability solutions
  • Exposure to diverse datasets across the supply chain
  • Collaboration with sustainability and data experts
  • Technical leadership role with strategic influence
  • Support for professional development in a cloud-focused environment
Full Job Description
Lead Data Software Engineer with Databricks We are seeking a Lead Data Software Engineer with Databricks expertise to drive the technical design, robustness, and evolution of end-to-end traceability solutions, enabling transparency, regulatory compliance, and data-driven decision-making across the supply chain. This role provides technical leadership and architectural direction for traceability platforms and data products, working with heterogeneous datasets ranging from farm and geolocation data to transactional, logistics, sustainability, and regulatory information. The successful candidate will collaborate closely with the Traceability solution team, sustainability experts, business analysts, data engineers, and data scientists, acting as a technical reference to ensure traceability data pipelines, data models, and platforms remain aligned with best practices and long-term architectural vision. Responsibilities Provide technical leadership and ownership for traceability data platforms, pipelines, and data products Define, document, and enforce architectural principles, technical standards, and best practices for traceability solutions Oversee the design of robust pipelines, quality controls, and proactive monitoring frameworks to ensure data availability, quality, integrity, and timeliness Guide and support data engineers and technical contributors on implementation choices, troubleshooting, performance optimization, and best practices Contribute strategically to Run execution and continuous improvement, including data quality, lineage, auditability, and operational monitoring Translate traceability and regulatory requirements into technical solutions, data models, and platform capabilities Collaborate closely with solution architects, data platform teams, and external partners to ensure alignment and scalability Support adoption of traceability solutions by business and sustainability teams through technical guidance and enablement Requirements 5+ years of experience in technical roles within data, analytics, or platform environments, with exposure to leadership or solution ownership responsibilities Expertise in cloud-based data platforms, preferably Azure Background in data architecture, data modeling, and data platform design, including complex, large-scale datasets Hands-on experience with Databricks, Spark/PySpark, and SQL for data transformation pipelines and analytics platforms Proficiency in SQL and data storage concepts, including relational databases, NoSQL, and data lakes, with performance optimization for large and complex datasets Knowledge of Python and data manipulation ecosystems such as Pandas and NumPy, with ability to review and guide code even if not hands-on full time Familiarity with data visualization tools including Power BI, Tableau, and Looker Experience designing or governing data quality frameworks, lineage, monitoring, and auditability Familiarity with orchestration tools and data integration patterns such as ADF, Airflow, and AWS Glue, or equivalent Experience with software engineering best practices including version control with git, branching strategies, code reviews, and CI/CD Understanding of data governance, security, and access control for sensitive business and regulatory data Proficiency in English at a B2+ level

About EPAM Systems

EPAM Systems, Inc. is a leading global provider of digital platform engineering and development services. The company has a strong presence in North America, Europe, and Asia, and serves clients in a variety of industries, including financial services, healthcare, and retail. EPAM's services include software engineering, product development, and digital platform engineering, and the company has a reputation for delivering high-quality solutions that help its clients achieve their business goals. EPAM has been recognized as a leader in the digital services industry by a number of independent research firms, and the company has won numerous awards for its work.
Learn more about EPAM Systems
Size
58,824 employees
Market Cap
$18.2 billion
Industry
Net Income
$327.1 million
Founded
1993
5 Year Trend
+26.5%
Revenue
$2.6 billion
NASDAQ

Similar Jobs

More Jobs at EPAM Systems

More Enterprise Technology Jobs

Find similar Lead Data Software Engineer with Databricks jobs: