Workday

Senior Software Engineer(Distributed Systems)

Workday$190K — $285K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 8+ years in software development engineering
  • 4+ years in designing and operating distributed systems
  • 5+ years experience with Java, Go, Scala, or Python
  • Bachelor’s degree in Computer Science or related field; Master's preferred
  • Strong skills in Algorithmic Thinking for scalable data solutions

Responsibilities

  • Own technical design for Workday's distributed tracing platform
  • Optimize low-latency infrastructure for multi-petabyte scale
  • Drive engineering excellence and mentor junior developers
  • Act as a domain expert for tracing ingestion and query performance
  • Contribute to end-to-end telemetry pipeline testing and CI/CD automation

Benefits

  • Comprehensive benefits package
  • Flexible work schedule with 50% in-office time per quarter
  • Opportunity for bonuses and stock grants
  • Collaborative team environment
  • Support for professional development and training opportunities
Full Job Description

About the Team

The Data Platform and Observability Engineering (DPOE) team is building Workday’s next-generation, multi-petabyte scale Observability Platform. We own the libraries, distributed services, and infrastructure that power ingestion, storage, and query across the observability stack — Iceberg, ClickHouse, Tempo, Grafana, S3, Kafka, and Elasticsearch serving traces, metrics, and logs for every workload at Workday. Our roadmap directly shapes how the company detects, diagnoses, and eventually predicts operational issues at scale.

About the Role

As a Senior Software Development Engineer, you will own the technical design and execution for core components of Workday’s distributed tracing platform, built on ClickHouse and/or Grafana Tempo and backed by a big-data pipeline (Kafka, Spark/Flink, Iceberg, S3). This is a hands-on role where you will tackle multi-petabyte scale challenges, optimize low-latency infrastructure, and help lay the technical groundwork for Observability AI.
You will drive engineering excellence within the team, mentor other developers, and act as a domain expert for tracing ingestion, storage, and query performance

About You

Basic Qualification 

8+ years experience in software development engineering.

4+ years experience specifically focused on designing, building, and operating complex distributed system architectures, evidenced by successful deployment of systems with high availability (e.g., 99.9% uptime) and fault tolerance.

5+ years experience with at least two of the following programming languages Java, Go, Scala, Python, including experience in writing production-level code for distributed systems.

Bachelor’s degree in a relevant field such as Computer Science, Engineering, or a related discipline; a Master's degree (e.g., MS in Computer Science, Distributed Systems, or related field) is strongly preferred or equivalent practical experience.

Other Qualification

Strong ability in Algorithmic Thinking, including to build highly efficient and scalable solutions for complex high-throughput data ingestion and sub-second query performance challenges.

Solid experience in API Development, including an understanding of gRPC, REST, and OpenTelemetry (OTLP), with practical experience designing and building scalable distributed APIs for observability data.

Strong understanding of Code Testing methodologies, such as distributed load testing and integration testing, and experience contributing to end-to-end telemetry pipeline testing and CI/CD automation.

Solid understanding of Distributed Systems Software principles, including data partitioning, eventual consistency, and fault tolerance mechanisms, with hands-on experience in Kafka, Spark, Flink, or ClickHouse.

Experience implementing and maintaining High Availability strategies for critical distributed systems, including multi-AZ deployments, robust retry mechanisms, and automated failover.

Practical experience with Large Scale Data Processing technologies and frameworks such as Apache Kafka, Spark, Flink, and Apache Iceberg within complex distributed architectures.

Good understanding of Large Scale Systems design principles, including distributed data sharding, replication, and query optimization, and experience working on observability pipelines or data lake platforms.

Strong understanding of Object-Oriented Design (OOD) principles and architectural patterns for building highly scalable and maintainable distributed systems.

Experience with Source Control Management (SCM) tools such as Git and GitHub/Bitbucket, and following best practices for collaborative distributed development workflows.

Strong understanding of System Security principles and best practices relevant to securing distributed environments, including mutual TLS (mTLS), multi-tenant authorization (authz), and data encryption.

Proven ability to actively collaborate within and across distributed software development teams and contribute constructively to architectural discussions and system designs.

Strong skills in creating Technical Writing Documentation for runbooks, system design specs, and API documentation related to distributed systems architecture and design.


Workday Pay Transparency Statement

The annualized base salary ranges for the primary location and any additional locations are listed below.  Workday pay ranges vary based on work location. As a part of the total compensation package, this role may be eligible for the Workday Bonus Plan or a role-specific commission/bonus, as well as annual refresh stock grants. Recruiters can share more detail during the hiring process. Each candidate’s compensation offer will be based on multiple factors including, but not limited to, geography, experience, skills, job duties, and business need, among other things. For more information regarding Workday’s comprehensive benefits, please .

Primary Location: USA.CA.PleasantonPrimary Location Base Pay Range: $190,100 USD - $285,100 USD


 

Additional US Location(s) Base Pay Range: $160,100 USD - $285,100 USD



Our Approach to Flexible Work
 

With Flex Work, we’re combining the best of both worlds: in-person time and remote. Our approach enables our teams to deepen connections, maintain a strong community, and do their best work. We know that flexibility can take shape in many ways, so rather than a number of required days in-office each week, we simply spend at least half (50%) of our time each quarter in the office or in the field with our customers, prospects, and partners (depending on role). This means you'll have the freedom to create a flexible schedule that caters to your business, team, and personal needs, while being intentional to make the most of time spent together. Those in our remote "home office" roles also have the opportunity to come together in our offices for important moments that matter.

About Workday

Workday, Inc. is a provider of enterprise cloud applications for finance and human resources. The Company delivers financial management, human capital management and analytics applications designed for various companies, educational institutions and government agencies. As part of its applications, the Company provides embedded analytics that capture the content and context of everyday business events, facilitating informed decision-making from wherever users are working. Its applications include Workday Financial Management, Workday Human Capital Management (HCM) and Other Applications. It also provides open, standards-based Web-services application programming interfaces, and pre-built packaged integrations and connectors. Workday, Inc. is headquartered in Pleasanton, California.
Learn more about Workday
Size
15,932 employees
Market Cap
$42.2 billion
Industry
Net Income
-$282.4 million
Founded
2005
5 Year Trend
+26.7%
Revenue
$4.3 billion
NASDAQ

Similar Jobs

More Jobs at Workday

More Information Technology Jobs

Find similar Senior Software Engineer(Distributed Systems) jobs: