Senior Site Reliability Engineer, NetBox Delivery

NetBox Labs

• $125K — $150K *
US-AnywhereRemote in United States
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years in software engineering, platform engineering, or SRE with maintainable code expertise
  • Production experience with Django and Postgres at scale
  • Strong skills in container builds, including security practices
  • Hands-on with AWS, Kubernetes, GitHub Actions, and related technologies
  • Experience building within an AI-augmented development framework
  • Proven ability to drive cross-team collaboration and migrations

Responsibilities

  • Own the build and release pipeline for NetBox
  • Ensure quick and predictable release handoffs to Cloud and Enterprise
  • Enhance application performance and reliability in production
  • Implement observability features for monitoring and alerting
  • Serve as an escalation point for performance issues and fixes
  • Strengthen supply chain security and support SOC 2 compliance
  • Lead incident response and postmortems for operational issues

Benefits

  • Opportunity to shape a new team within the organization
  • Work with a diverse technology stack
  • Engage with both open source and commercial software aspects
  • Be part of a culture of continuous improvement and innovation
  • Collaboration with cross-functional teams
Full Job Description
NetBox Labs is seeking a Senior Site Reliability Engineer for NetBox Delivery, a new team in our Applications group.

NetBox is a product that reaches people in a few ways: our open source community runs NetBox OSS; commercial customers use NetBox Cloud (SaaS) or NetBox Enterprise (self-managed). NetBox Delivery owns everything between a NetBox Core release and a healthy, running instance on Cloud and Enterprise. We ship the software, keep an eye on it in production, and when something breaks, we fix it at the source rather than working around it. You will be one of the first engineers on this team and help shape how it works.

In this role you will:

  • Own the NetBox build and release pipeline, from base images to downstream availability on Cloud and Enterprise
  • Build the release handoff between NetBox Core and the Cloud and Enterprise teams, so new releases reach customers quickly and predictably
  • Make NetBox faster and more reliable in production, from application startup to Postgres performance
  • Build real observability for the application and the release pipeline, including monitoring, alerting, and SLOs
  • Act as the escalation point for performance and reliability issues and take fixes back to NetBox Core when the cause is in the code
  • Strengthen supply chain security and support SOC 2 compliance for the build pipeline
  • Share on-call duties and lead incident response and postmortems for your area


Requirements:
  • 5+ years in software engineering, platform engineering, or SRE, with proven experience writing robust, maintainable code
  • Production experience with Django and Postgres at scale, including schema design, migration risk, and query performance under real load
  • Strong container build skills, including base image design, Python dependency management, and supply chain security practices like vulnerability scanning and image signing
  • Hands-on experience with our stack or something close to it: AWS (EC2, VPC, IAM, RDS), Kubernetes and Helm, GitHub Actions, ArgoCD or FluxCD, Terraform, and Prometheus and Grafana
  • Hands-on experience building inside an AI-augmented development harness, including Claude Code and the workflows that make agentic tooling reliable
  • A track record of driving work across team boundaries, from writing the RFC to getting a cross-team migration done


Nice to haves:
  • Familiarity with the NetBox ecosystem or network automation
  • Open source contributions or project involvement
  • Experience working in a B2B software startup or high-growth organization
  • Deep experience with supply chain security tooling such as cosign, Sigstore, or SLSA
  • Experience operating high-throughput or performance-sensitive systems for large enterprise customers


Similar Jobs

More Jobs at NetBox Labs

More Information Technology Jobs

Find similar Senior Site Reliability Engineer, NetBox Delivery jobs: