Senior/Staff DevOps Engineer, Platform Infrastructure

Trase Systems

• $180K — $240K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 10+ years of experience in software, platform, infrastructure, SRE, or DevOps engineering with production systems ownership.
  • Deep hands-on experience with Linux, Docker, Kubernetes, and production cluster operations.
  • Strong experience with Helm and IaC tools like Terraform or Pulumi, including reusable modules and change review.
  • Hands-on experience with at least two among AWS, Azure, and GCP, with knowledge of core services.
  • Experience operating software in customer-controlled, private-cloud, hybrid-cloud, or on-premises environments.
  • Strong understanding of networking and security concepts related to cloud deployments.
  • Experience with CI/CD systems including monitoring and incident response metrics.

Responsibilities

  • Architect, build, and operate secure multi-cloud infrastructure for Trase OS.
  • Containerize application services using Docker, Kubernetes, Helm, and similar tools.
  • Create reusable infrastructure-as-code modules for deployment workflows.
  • Design deployment patterns catering to varied customer-specific requirements.
  • Abstract hard dependencies on cloud services for improved portability.
  • Troubleshoot complex issues across various environments and infrastructure layers.
  • Build observability frameworks for enhanced diagnostics and support.

Benefits

  • Career advancement opportunities aligned with company growth.
  • Comprehensive health care coverage with 100% employer contribution for employees and families.
  • 14 weeks paid maternity and paternity leave at regular pay.
  • Unlimited PTO subject to management approval.
  • Professional development opportunities for continued learning.
  • Optional benefits including 401K, FSA, and equity incentives.
  • Mental health benefits available through Tara Mind.
Full Job Description
About the Role

As a Senior or Staff DevOps Engineer on Platform Infrastructure, you will design and operate the infrastructure that supports Trase OS across Trase-hosted and customer-controlled environments. You will build secure, repeatable deployment patterns spanning AWS, Microsoft Azure, Google Cloud Platform, private cloud, hybrid cloud, and on-premises infrastructure.

This is a hands-on engineering role with broad ownership. You will work across application packaging, Kubernetes, infrastructure as code (IaC), networking, security, observability, release engineering, and production reliability. You will also partner directly with engineering and customer-facing teams to turn deployment requirements into systems that can be installed, upgraded, operated, and supported consistently.

The level will reflect your experience and demonstrated scope. Staff-level candidates will be expected to lead architecture across teams, establish engineering standards, and mentor other engineers.
Why this Role is Needed

Trase OS supports mission-critical, long-running workflows in environments with different cloud services, network controls, security requirements, and operating models. Our deployment architecture must remain portable without sacrificing reliability, security, or operational clarity.

This role will reduce one-off deployment work, remove avoidable dependencies on a single cloud provider, and establish reusable infrastructure that internal teams and customers can operate with confidence.
What You'll Do
  • Architect, build, and operate secure infrastructure across AWS, Microsoft Azure, and Google Cloud Platform, as well as private-cloud, hybrid-cloud, on-premises, and customer-controlled environments.
  • Containerize and package Trase OS application services using Docker, Kubernetes, Helm, Kustomize, or equivalent tools.
  • Create reusable infrastructure-as-code (IaC) modules and deployment workflows using Terraform, Pulumi, or comparable tooling.
  • Design deployment patterns that account for customer-specific requirements such as restricted networks, limited or no egress, approved registries, data residency, and cloud account ownership.
  • Identify, replace, or abstract hard dependencies on managed cloud services when they prevent portability across deployment environments.
  • Troubleshoot complex issues across applications, Kubernetes clusters, cloud services, networks, and infrastructure rather than treating platform work as CI/CD scripting alone.
  • Build and maintain CI/CD and GitOps workflows, release orchestration, environment promotion, upgrade paths, rollback procedures, and version compatibility controls.
  • Build and operate observability for Trase OS across metrics, logs, traces, dashboards, alerting, and SLOs so teams can diagnose failures and support customer deployments.
Technical Leadership
  • Lead infrastructure and reliability decisions that affect multiple product and customer teams.
  • Set practical standards for production readiness, cloud portability, security, observability, and operational support.
  • Make clear tradeoffs between speed, maintainability, cost, security, and customer requirements, and drive decisions through implementation.
  • Build reusable platform capabilities that replace one-off deployment solutions.
  • Mentor engineers and raise the team's ability to design, ship, and operate distributed systems.
Qualifications
  • 10+ years of software, platform, infrastructure, SRE, or DevOps engineering experience, including ownership of production systems.
  • Deep hands-on experience with Linux, Docker, Kubernetes, and production cluster operations.
  • Strong experience with Helm and infrastructure as code, such as Terraform or Pulumi, including reusable modules, state management, testing, and change review.
  • Hands-on infrastructure experience with at least two of AWS, Azure, and GCP, with working knowledge of core compute, networking, storage, identity, and managed-service patterns across all three.
  • Experience deploying and operating software across customer-controlled, private-cloud, hybrid-cloud, or on-premises environments, including adapting cloud-native SaaS products for these deployment models.
  • Strong understanding of networking, DNS, ingress, load balancing, certificates, IAM, secrets management, persistent storage, and service-to-service security.
  • Experience building and operating CI/CD or GitOps systems, including production releases, upgrades, and rollbacks, with metrics, logs, traces, SLOs, and alerts for incident response.
  • Strong software engineering and automation skills in Python, Go, TypeScript, or a similar language, with the ability to work across application and infrastructure layers.
  • Experience designing secure, portable production infrastructure, including hardening systems and replacing or abstracting cloud-specific managed services when needed.
  • Experience converting a cloud-native SaaS product into a customer-hosted, private-cloud, or on-premises deployment model.
  • Demonstrated experience using AI-assisted coding and engineering tools to accelerate development, infrastructure automation, troubleshooting, operational analysis, or incident investigation.
Preferred
  • Experience operating in restricted-network, disconnected, regulated, or security-sensitive environments.
  • Familiarity with compliance and security frameworks such as HIPAA, SOC 2, NIST, FedRAMP, or related government requirements.
  • Experience with service mesh, policy as code, admission controls, software supply-chain security, artifact signing, or software bills of materials.
  • Hands-on experience with Crossplane or meaningful contributions to CNCF projects, such as code, documentation, design, or community maintenance.
  • Experience supporting long-running, stateful, data-intensive, AI/ML, or GPU-enabled workloads.
  • Customer-facing engineering, solutions architecture, or forward-deployed engineering experience.

Up to 20% travel may be required.

If you want to be on the cutting edge of technology, building AI solutions for the future, and are up for a challenge, let's talk!

Salary Range: $180,000-$240,000 + bonus + equity. This represents the typical salary range for this position based on experience, skills, and other factors.

Our Trase Benefits:

For full-time roles only
  • Career track opportunity with potential for rapid advancement with strong performance as the firm grows
  • 100% employer paid, comprehensive health care including medical, dental, and vision for you and your family.
  • Paid maternity and paternity for 14 weeks at employees' normal pay.
  • Unlimited PTO, with management approval.
  • Opportunities for professional development and continued learning.
  • Optional 401K, FSA, and equity incentives available.
  • Mental health benefits are available through Tara Mind.

Similar Jobs

More Jobs at Trase Systems

More Information Technology Jobs

Find similar Senior/Staff DevOps Engineer, Platform Infrastructure jobs: