The Aerospace Corporation

Kubernetes Site Reliability Engineer

The Aerospace Corporation$129K — $193K *
Aerospace & Defense
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in STEM or related fields.
  • 8+ years of experience in large-scale software systems.
  • 5+ years supporting highly available enterprise environments.
  • 2+ years managing Kubernetes environments.
  • Experience with AWS and Azure deployment and management.
  • Strong Linux system administration skills.
  • Ability to independently resolve engineering problems.

Responsibilities

  • Develop and maintain Kubernetes-based Platform as a Service services.
  • Manage Kubernetes production environments in on-premises and cloud setups.
  • Perform full security patching of Kubernetes infrastructure.
  • Own engineering responsibilities for production AWS and Kubernetes services.
  • Identify and solve Kubernetes stack engineering issues.
  • Ensure real-time telemetry data analysis capabilities.
  • Provide support during launch events as needed.

Benefits

  • Comprehensive health care and wellness plans.
  • Paid holidays, sick time, and vacation.
  • Flexible work schedules and telework options.
Full Job Description

The Digital Innovation Division (DID) is accountable for integrating strategies, providing governance, and managing internal investments that form the foundation of Aerospace’sdigital innovation and transformation. The DID Mission IT pillar supports engineering teams across Aerospace by deliveringtop-tier IT engineering and IT support servicestailored to meet the unique needs of our engineering community.

Mission IT Operationsis seeking a skilledSite ReliabilityEngineerwithdeepexpertisein Kubernetes, Linux, programming, and automation. In this role, you willbe responsible fordeveloping andmaintainingboth on-premisesand cloud-based Kubernetes clustersthat form the core of an overall Platform as a Service (PaaS), providing essential support to our engineering team.

As part of a multidisciplinaryplatform andinfrastructure team, you will manage multipleKubernetesclusters used for technical analysessuchspace launch telemetryanalysisand modelingand simulation, as well as for Artificial Intelligence (AI) Large Language Model(LLM)Training and inference services.Collaborating closely with rocket scientists and engineers, you will contribute to the development of innovative solutions to complex challenges within the space enterprise, supporting critical national space assets. This position requiresa strong senseof shared responsibility and ownership, working alongside cross-functional team members to achieve our missionobjectives.

Work Model:This is a full-time position based in El Segundo, CA whichrequires100% onsite work.

WhatYoullBe Doing

  • Developing and sustaining advanced services for our Kubernetes-based PaaS (e.g., Coder workspaces,Kueuebatching scheduling,Knativeserverless,Crossplanecontrol planes)

  • ManagingKubernetes forproductionon-premisesandcloudenvironments(e.g., AWS, Azure)with end-to-end responsibilities of deployment, upgrade, patching, performance tuning, capacity planning,andbackups/DR.

  • Frequent full security patching of all layers of Kubernetes infrastructure whilemaintainingvery highuptime

  • Ownership and engineering responsibility of production AWS and Kubernetes services

  • Identifyingand resolvingfull Kubernetes stack engineering problems independently

  • Ensuringsuccessful real-time analysis of telemetry data from space launch partners, such as SpaceX, United Launch Alliance (ULA), and Blue Origin

  • Providingafter-hours support for Kubernetes infrastructure troubleshooting during launch events

  • Supportingscientists and engineers runningapplicationsin Kubernetes

  • ProvidingLinuxexpertiseand troubleshooting

  • Evaluatingand testingnew products and technologies

  • Usingcode to enhance and automate operations

What You Need to be Successful

MinimumRequirements forEngineering Specialist:

  • Bachelors degree in STEM, Computer Science.or other related sciences/engineering discipline.

  • 8 or more years of relevant experiencedirectly relatedto developing and delivering complex large-scale distributed software systems solutions and technical products

  • Minimum of 5 years experience supporting highly available enterprise environments, including maintaining system uptime and service availability targets.

  • At least 2 yearsofhands-on experience managing existing Kubernetes environments, with responsibilities of deployment, upgrade, patching, and backups

  • Full ownership and engineering responsibility of production Kubernetes services, both on-premises and Cloud Service Providers such as AWS and Azure

  • Ability toidentifyand resolve engineering problems independently

  • Experience in Linux systems administration, including configuration, for an enterprise environment

  • Strong understanding of networkingand storagefundamentals

  • Experience automating repetitive tasks with scripting or DevOps tools

  • This position requires the ability to obtain a TS/SCI security clearance and polygraph, which is issued by the U.S. government. U.S. citizenship isrequiredto obtain a security clearance.

In addition to the above, theminimumrequirements forSenior Engineering Specialistinclude:

  • 12 or more years of relevant experiencedirectly relatedto developing and delivering complex large-scale distributed software systems solutions and technical products

  • 8yearsofexperience supportinga highly availableenterprise environment

  • Experience architecting and deploying securecloud (e.g., AWS, Azure)and/or Kubernetes environments from scratch

  • Experience performance tuning and capacity planningcloud (e.g., AWS, Azure)and/or Kubernetes environments

How You can Stand Out

It would be impressive if you have one or more of these:

  • A current and active U.S. GovernmentTS/SCI security clearanceand polygraph

  • Certified Kubernetes Administrator (CKA), Red Hat Certified System Administrator (RHCSA), or Red Hat Certified Engineer (RHCE)

  • ExperiencemanagingKubernetes clusters usingRancher

  • Experience deploying/supportingpersistent container storage on Kubernetes(i.e.,Portworx, RookCeph,OpenEBS, Longhorn)

  • Experience inLinux performance tuning, and security hardening for an enterpriseDoWenvironment

  • Experience withVMwareor Harvestervirtualization infrastructures

  • Experience with automated provisioning, configuration management, Infrastructure-as-Code,GitOps(i.e.,ArgoCD, Ansible,TerraForm, Puppet, Packer, Bash, Golang, Python)

  • Experience with Agile and Scrum

We offer a competitive compensation package where youll be rewarded based on your performance and recognized for the value you bring to our business. The grade-based pay range for this job is listed below. Individual salaries within that range are determined through a wide variety of factors including but not limited to education, experience, knowledge and skills.

(Min - Max)

$129,000.00 - $193,500.00

Pay Basis: Annual

Leadership Competencies

Our leadership philosophy is simple: every employee, regardless of level and role, can demonstrate leadership. At Aerospace, our commitment is our people. To cultivate our talent and ensure that we have a strong pipeline of future leaders, we want individuals who:

  • Operate Strategically
  • Lead Change
  • Engage with Impact
  • Foster Innovation
  • Deliver Results

Ways We Reward Our Employees

During your interview process, our team will provide details of our industry-leading benefits.

Benefits vary and are applicable based on Job Type. A few highlights include:

  • Comprehensive health care and wellness plans

  • Paid holidays, sick time, and vacation

  • Standard and alternate work schedules, including telework options<

About The Aerospace Corporation

The Aerospace Corporation is a nonprofit corporation that operates a federally funded research and development center (FFRDC) headquartered in El Segundo, California. The corporation provides technical guidance and advice on all aspects of space missions to military, civil, and commercial customers. Aerospace also designs and develops spacecraft, sensors, and other systems in support of national security, civil, and commercial customers. The corporation has more than 4,000 employees and operates a number of laboratories and test facilities across the United States.
Learn more about The Aerospace Corporation
Size
4,000 employees
Industry
Founded
1960

Similar Jobs

More Jobs at The Aerospace Corporation

More Aerospace & Defense Jobs

Find similar Kubernetes Site Reliability Engineer jobs: