Skydio

Site Reliability Engineer

Skydio • $180K — $220K *
US-AnywhereRemote in San Mateo, CA
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 3+ years of experience in Production Engineering, SRE, DevOps, or a similar role.
  • Strong hands-on experience with Kubernetes, beyond just application deployment.
  • Experience managing Kubernetes/EKS upgrades and production clusters.
  • Solid understanding of AWS fundamentals, including VPCs and IAM.
  • Proficient in Terraform or similar infrastructure-as-code tools.
  • Experience with CI/CD systems like Argo CD or Jenkins.
  • Ability to diagnose production infrastructure and networking issues.

Responsibilities

  • Build, operate, and troubleshoot production Kubernetes/EKS clusters.
  • Perform upgrades and maintenance on Kubernetes clusters.
  • Manage AWS infrastructure components like VPCs and load balancers.
  • Define and maintain infrastructure using Terraform.
  • Build and operate CI/CD and deployment infrastructure.
  • Troubleshoot production issues across various platforms.
  • Implement monitoring and observability for critical infrastructure.

Benefits

  • Paid vacation time and sick leave.
  • Holiday pay and 401K savings plan.
  • Comprehensive health insurance plans.
  • Equity in the form of stock options.
Full Job Description
About the role:

We are looking for a hands-on Site Reliability Engineer to build, operate, and scale the cloud infrastructure that powers our products. This role is focused on owning production infrastructure, including Kubernetes, AWS, infrastructure as code, CI/CD, observability, networking, and reliability.

You don't need to be an expert in every area, but you should have strong Kubernetes and cloud fundamentals with meaningful depth in at least one infrastructure domain. Our technology helps save lives. You'll play a critical role in keeping the infrastructure behind it reliable, scalable, and available when it matters most.

How you'll make an impact:
  • Build, operate, and troubleshoot production Kubernetes/EKS clusters.
  • Perform Kubernetes upgrades, node rollouts, and cluster maintenance.
  • Build and manage AWS infrastructure including VPCs, networking, subnets, load balancers, IAM, EKS, databases, and storage.
  • Define and maintain infrastructure using Terraform.
  • Build and operate CI/CD and deployment infrastructure.
  • Troubleshoot production issues across Kubernetes, AWS, Linux, networking, and databases.
  • Build monitoring, alerting, and observability for critical infrastructure.
  • Participate in on-call rotations and respond to production incidents.
  • Identify and solve infrastructure scaling and reliability problems.
  • Automate operational work using Python, Go, or similar languages.
  • Help expand infrastructure across new regions and deployment environments.

What makes you a good fit:
  • 3+ years of experience as a Production Engineer, SRE, DevOps, or equivalent infrastructure role.
  • Strong hands-on experience operating Kubernetes, not simply deploying applications to existing clusters.
  • Experience managing Kubernetes/EKS upgrades and production clusters.
  • Strong AWS fundamentals, including VPCs, public/private subnets, networking, load balancers, EKS, IAM, and databases.
  • Production experience with Terraform or similar infrastructure-as-code tooling.
  • Experience owning or maintaining CI/CD and deployment systems such as Argo CD, Spinnaker, GitHub Actions, GitLab CI/CD, or Jenkins.
  • Experience diagnosing production infrastructure and networking problems.
  • Experience solving meaningful scaling or reliability challenges.
  • Obtaining FAA Part 107 certification within the first 60 days of employment is strongly encouraged for all Skydio employees and required for certain positions.
  • This position requires access to export-controlled technical data, restricted government information, and/or information systems subject to U.S. government security and access-control requirements. Employment in this role is contingent upon verification of U.S. person status and the ability to access controlled or restricted information as required for the position.

Bonus points:
  • Helm and GitOps experience.
  • Datadog or similar observability tooling.
  • PostgreSQL/database operations experience.
  • Multi-region infrastructure experience.
  • On-premises or disconnected deployment experience.
  • Streaming or high-throughput distributed systems experience.

Compensation:

At Skydio, our compensation packages for regular, full-time employees include competitive base salaries, equity in the form of stock options, and comprehensive benefits packages. Compensation will vary based on factors, including skill level, proficiencies, transferable knowledge, and experience. Relocation assistance may also be provided for eligible roles. The annual base salary range for this position is $180,000 - $220,000*. Fundamentally, we believe that equity is the key to long-term financial growth, and we ensure all regular, full-time employees have the opportunity to significantly benefit from the company's success. Regular, full-time employees are eligible to enroll in the Company's group health insurance plans. Regular, full-time employees are eligible to receive the following benefits: Paid vacation time, sick leave, holiday pay and 401K savings plan. This position and all associated benefits are subject to applicable federal, state, and local laws, as well as the Company's policies and eligibility criteria.

*Compensation for certain positions may vary based on the position's location.

#LI-WA1

About Skydio

Skydio is a leading manufacturer of autonomous drones for consumer and commercial use. The company's drones are equipped with advanced computer vision and artificial intelligence technology, allowing them to navigate complex environments and avoid obstacles. Skydio was founded in 2014 by a team of experts in robotics, computer vision, and artificial intelligence. The company is headquartered in Sunnyvale, California.
Learn more about Skydio
Size
200 employees
Industry
Founded
2014

Similar Jobs

More Jobs at Skydio

More Information Technology Jobs

Find similar Site Reliability Engineer jobs: