Domino Data Lab

Staff Performance Quality Engineer

Domino Data Lab$185K — $210K *
US-AnywhereRemote in United States
Enterprise Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Experience with complex automation platforms (performance/load testing, end-to-end frameworks, CI systems)
  • Strong proficiency in Python, especially within automation-heavy codebases
  • Hands-on experience in distributed systems and reliability engineering
  • Knowledge of cloud-based infrastructure scaling, particularly using Kubernetes
  • Proven ability to design multi-stage test automation systems
  • Ownership mindset with self-direction and effective communication

Responsibilities

  • Serve as the technical owner of Tempest, ensuring its reliability and alignment with product needs
  • Modernize scale testing infrastructure for growth and product complexity
  • Provide accurate, data-driven size recommendations for documentation via empirical testing
  • Enhance validation and reporting in scale testing through automated criteria
  • Operationalize scale testing on cloud platforms with appropriate sizing guidance
  • Collaborate with platform teams for effective scale testing across diverse cloud providers
  • Build automation to increase team efficiency as product and customer base grows

Benefits

  • Equity opportunities
  • Company bonus or sales commission potential
  • 401(k) plan
  • Medical, dental, and vision benefits
  • Wellness stipends
Full Job Description
What we are building

The Automation Team at Domino acts as a force multiplier for engineering, building the tools and systems that enable teams to ship code confidently and consistently. A core part of this mission is Tempest, an in-house platform that orchestrates realistic, long-duration workloads against live Kubernetes clusters and validates the results against real observability data. Today, when scale testing surfaces a bottleneck, a resource misconfiguration, or a regression in system behavior, the team can identify and report the issue - but we need someone who can take the next step: profiling services, tracing root causes through Prometheus and New Relic data, and partnering with platform engineers to drive durable fixes. Focused on iteration and continuous improvement, the team looks for targeted enhancements that create outsized impact, and this role will close the gap between detection and resolution at the infrastructure level.

What your impact will be

In your first year, you will:
  • Serve as the technical owner of Tempest, Domino's scale and reliability platform, ensuring it remains reliable, extensible, and aligned with evolving infrastructure needs
  • Diagnose and drive resolution of performance bottlenecks and resource misconfigurations surfaced by scale testing - working directly with platform and infrastructure teams to ship fixes, not just file tickets
  • Deliver accurate, data-driven sizing recommendations for customer-facing documentation based on rigorous empirical testing across deployment sizes
  • Strengthen observability across scale testing by improving Prometheus and New Relic instrumentation, making it faster to pinpoint root causes during and after multi-day load runs
  • Establish and operationalize scale testing on cloud platforms, ensuring appropriate sizing and configuration guidance for this increasingly divergent product line
  • Partner with platform teams to enable effective scale and reliability testing across additional cloud providers, helping position Domino for future multi-cloud success
  • Increase the efficiency and leverage of a small team by building infrastructure automation that scales operationally as the product and customer base grow

What we look for in this role
  • Background in SRE, platform engineering, or infrastructure with hands-on experience operating and troubleshooting distributed systems in production Kubernetes environments
  • Strong proficiency in Python and comfort working in a large, modular codebase that spans orchestration, infrastructure automation, and systems integration
  • Experience with observability stacks (Prometheus, Grafana, New Relic, or similar) - writing queries, building dashboards, and using metrics to diagnose performance and reliability issues at the systems level
  • Demonstrated ability to go beyond detection to resolution: profiling services, identifying resource bottlenecks, and working with engineering teams to ship durable fixes
  • Familiarity with performance and load testing methodologies (e.g., Locust, k6, or similar) as part of a broader infrastructure or reliability practice
  • Clear ownership mindset - self-directed, accountable, and able to communicate priorities and status effectively in a remote, async environment

What we value
  • We value a growth mindset. High-performing creative individuals who dig into problems and see the opportunities for success
  • We believe in individuals who seek truth and speak the truth and can be their whole selves at work
  • We value all of you that believe improving is always possible At Domino Everything is a work in progress - we can do better at everything
  • We emphasize an environment of teaching and learning to equip employees with the tools needed to be successful in their function and the company
  • We strongly believe in the value of growing a diverse team and encourage people of all backgrounds, genders, ethnicities, abilities, and sexual orientations to apply

#LI-Remote

The annual US base salary range for this role is listed below. For sales roles, the range provided is the role's On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. This salary range will be narrowed during the interview process based on a number of factors, including the candidate's experience, qualifications, and location. Additional benefits for this role may include: equity, company bonus or sales commissions/bonuses; 401(k) plan; medical, dental, and vision benefits; and wellness stipends.

Compensation Range

$185,000-$210,000 USD

About Domino Data Lab

Domino Data Lab is a software company that provides a platform for data science teams to collaborate and build models. The company was founded in 2013 and is headquartered in Oakland, California. Domino Data Lab's platform allows data scientists to work together on projects, share code and data, and track experiments. The company also offers tools for model deployment and management. Domino Data Lab aims to help organizations make better decisions by leveraging the power of data science.
Learn more about Domino Data Lab
Size
200 employees
Industry
Founded
2013

Similar Jobs

More Jobs at Domino Data Lab

More Enterprise Technology Jobs

Find similar Staff Performance Quality Engineer jobs: