Sonatype

Staff Data Engineer

Sonatype$135K — $160K *
US-AnywhereRemote in Canada
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 8+ years of experience in Data Engineering or a related backend engineering role
  • Bachelor's degree in Computer Science, Engineering, or a similar technical field
  • Expertise in tuning Spark jobs and managing Delta Lake architecture
  • Experience with AI-assisted development tools and AI/ML technologies
  • Strong programming skills in Python, Scala, or Java
  • Experience with distributed data systems like Spark or Kafka
  • Proficient in writing and optimizing complex SQL and NoSQL queries

Responsibilities

  • Design, build, and maintain scalable data pipelines and ETL/ELT processes
  • Architect and optimize data models and storage solutions for different use cases
  • Collaborate with data scientists, analysts, and engineers to ensure data quality
  • Own and evolve parts of the data platform using Databricks and Spark
  • Implement observability and data quality monitoring for critical pipelines
  • Drive best practices in data engineering including CI/CD and documentation
  • Contribute to the design of the next-generation data lakehouse architecture

Benefits

  • Parental leave policy
  • Paid volunteer time off (VTO)
  • Diversity and inclusion working groups
  • Flexible working practices
  • Recognition as a leader in software supply chain security and diversity initiatives
Full Job Description
We\'re looking for a Staff Data Engineer to join our growing Data Platform team. You\'ll play a key role in designing and scaling the infrastructure and pipelines that power analytics, machine learning, and business intelligence across Sonatype.You\'ll work closely with stakeholders across product, engineering, and business teams to ensure data is reliable, accessible, and actionable. This role is ideal for someone who thrives on solving complex data challenges at scale and enjoys building high-quality, maintainable systems.
At Sonatype, we:
  • Use data with purpose: you\'ll get the chance to work on problems that directly impact how the world builds secure software
  • Use modern tooling: you\'ll get the chance to leverage the best of open-source and cloud-native technologies
  • Have a deep collaborative culture: you\'ll be joining a passionate team that values learning, autonomy, and impact


What you\'ll do:

  • Design, build, and maintain scalable data pipelines and ETL/ELT processes
  • Architect and optimize data models and storage solutions for analytics and operational use
  • Collaborate with data scientists, analysts, and engineers to deliver trusted, high-quality datasets
  • Own and evolve parts of our data platform using Databricks and Spark
  • Implement observability, alerting, and data quality monitoring for critical pipelines
  • Drive best practices in data engineering, including documentation, testing, and CI/CD
  • As a Staff Engineer you will help drive long-term architectural vision and mentor the team on engineering best practices, while partnering with stakeholders to ensure data solutions support business outcomes.
  • Contribute to the design and evolution of our next-generation data lakehouse architecture


What you bring:

  • 8+ years of experience as a Data Engineer or in a similar backend engineering role
  • Bachelor\'s degree in Computer Science, Engineering, or a related technical field
  • Databricks Optimization: Tune Spark jobs, optimize join performance, and manage Delta Lake architecture for batch and streaming data.
  • Experience leveraging AI-assisted development tools and AI/ML technologies to improve data engineering workflows, developer productivity, data quality and ops.
  • Strong programming skills in Python, Scala, or Java
  • Hands-on experience with distributed data systems like Spark or Kafka
  • Proficient in writing complex SQL and NoSQL queries and optimizing queries for performance
  • Experience building and maintaining robust ETL/ELT pipelines in production
  • Understanding of data modeling techniques (star schema, dimensional modeling, etc.)


It\'d be great if you also had:

  • Familiarity with software supply chain, cybersecurity, or large-scale software ecosystem data
  • A track record of improving data platform reliability, scalability, performance, and cost efficiency
  • Familiarity with workflow orchestration tools (Airflow, Dagster, or similar)
  • Hands-on experience with cloud data platforms, particularly AWS
  • Familiarity with modern table formats such as Delta Lake, Apache Iceberg, or Apache Hudi
  • Experience implementing data observability, lineage, governance, and automated data quality frameworks
  • Experience designing real-time or streaming data architectures using data lake technologies


Things we are proud of:

  • 2026 Gartner® Magic Quadrant™ Leader for Software Supply Chain Security
  • 2026 Celebrating 15 Years of Sonatype Research Labs - Industry-leading software supply chain and open source security research
  • 2026 Founding Member of the Linux Foundation Initiative for Open Source Sustainability
  • 2026 State of the Software Supply Chain® Report - Continuing industry leadership in software supply chain security and AI security research
  • 2025 Visionary in Gartner® Magic Quadrant™ for Application Security Testing!
  • 2025 AI Compliance Solution of the Year - AI Breakthrough Awards
  • 2025 DEVIES Award to our SBOM Manager for a new product for its innovation and impact in developer technology
  • 2024 Industry Leader in Forrester-Wave for Software Composition Analysis (2024 Q4 report)
  • Constellation AST Shortlist: Sonatype has been listed on the Constellation ShortList™ for Application Security Testing for 2024
  • Data Breakthrough Awards: Sonatype was announced as a 2024 winner in the \"Open Source Data Solution of the Year.\"
  • SD Times: Best in Show Security
  • Fast Company Best Workplaces for Innovators 2024
  • The Herd Top 100 Private Software Companies 2024
  • Diversity & Inclusion Working Groups
  • Parental Leave Policy
  • Paid Volunteer Time Off (VTO)


At Sonatype, we value diversity and inclusivity. We offer perks such as parental leave, diversity and inclusion working groups, and flexible working practices to allow our employees to show up as their whole selves.

About Sonatype

Sonatype is a software company that provides open-source governance and DevSecOps automation tools. The company was founded in 2008 by Jason van Zyl and Wayne Jackson and is headquartered in Fulton, Maryland. Sonatype's products include Nexus Repository, Nexus Lifecycle, and Nexus Firewall, which are used by organizations to manage and secure their software supply chains. The company's clients include Fortune 500 companies, government agencies, and leading software companies. Sonatype has received several awards for its innovative products and has been recognized as a leader in the software industry.
Learn more about Sonatype
Size
500 employees
Industry
Founded
2008

Similar Jobs

More Jobs at Sonatype

More Information Technology Jobs

Find similar Staff Data Engineer jobs: