Okta

Staff TDI Site Reliability Engineer, Okta Federal

Okta$174K — $239K *
Aerospace & Defense
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 7+ years of experience in SRE, DevOps, Cloud Automation, or Systems Engineering delivering complex infrastructure projects.
  • Proficient in container orchestration environments like EKS, ECS Fargate, and general container usage.
  • Skilled in infrastructure automation using Terraform and Python with a focus on secure software development.
  • Familiar with monitoring tools, particularly Splunk, CloudWatch, and Grafana stack.
  • Knowledge of networking concepts including BGP and IPsec management, utilizing AWS networking services.
  • Active U.S. TS/SCI security clearance with polygraph.

Responsibilities

  • Operate and maintain enterprise-grade solutions in air-gapped environments.
  • Build, run, and monitor secure development tools and infrastructure.
  • Autonomously manage operations within secure facilities.
  • Maintain SLOs/SLIs for workloads without external monitoring dependencies.
  • Own and manage runbooks and incident response procedures for limited escalation paths.
  • Participate in POA&M remediation and support Authority to Operate activities.
  • Support and run mission-critical services for product teams.

Benefits

  • Health, dental, and vision insurance.
  • 401(k) plan with company match.
  • Flexible spending accounts.
  • Generous paid leave policy including PTO and parental leave.
Full Job Description
The Staff Site Reliability Engineer OpportunityOkta Federal, Inc. is looking for an experienced Staff TDI Site Reliability Engineer to help build, improve, and maintain our cloud platform services that help Okta support the most sensitive national security missions. The Site Reliability Engineering team delivers foundational infrastructure capabilities that enable corporate engineering teams to operate securely, reliably, and at scale. You'll play a key role in designing and implementing complex cloud-based engineering enablement systems, while ensuring compliance with strict government requirements.

What you'll be doing
  • Operate and maintain enterprise grade solutions within air-gapped environments.
  • Build, run, and monitor development tools, pipelines, and infrastructure with a security-first mindset.
  • Operate autonomously within secure facilities.
  • Maintain SLOs/SLIs for workloads with no dependency on external monitoring or SaaS tooling.
  • Own runbooks and incident response procedures tailored to limited external escalation paths.
  • Participate in POA&M remediation and support annual/recurring Authority to Operate activities.
  • Support and run mission critical services depended on by product teams
  • Deliver excellent internal customer service and advocate for SRE and DevOps practices across teams.
  • Build and operate CI/CD pipelines that function without internet connectivity.

What you'll bring to the role
  • 7+ years of experience as an SRE, DevOps Engineer, Cloud Automation Engineer, or Systems Engineer with a track record of delivering complex infrastructure projects at scale.
  • Experience with container orchestration and runtime environments, including EKS, ECS Fargate, and general container usage.
  • Proficient in infrastructure automation using Terraform and developing automation tools with Python, while leveraging secure software development practices.
  • Experience with monitoring tools, especially Splunk, CloudWatch, and the Grafana stack.
  • Experience with general networking concepts, such as BGP and IPsec management, and has leveraged AWS networking services, including VPCs, TGWs, and VPC endpoints.
  • Security Clearance: Active U.S. TS/SCI with polygraph.

Additional requirements
  • The selected candidate may be subject to drug testing to the extent required by U.S. Government contracts.

And extra credit if you have experience in the following
  • Knowledgeable in Linux system administration.
  • Experience in secure and compliant environments (e.g., FedRAMP), with understanding of FIPS, STIG, and data boundary implementations.
  • Working experience operating tools and services in air gapped environments.



#LI-hybrid

Below is the annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York and Washington. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: https://rewards.okta.com/us.

The annual base salary range for this position for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York, and Washington is between:

$174,000-$239,000 USD

About Okta

Okta is a leading provider of identity and access management solutions for enterprises. The company's cloud-based platform enables organizations to securely connect people and technology, providing secure access to applications and data from any device, anywhere, at any time. Okta's solutions are used by thousands of organizations worldwide, including many Fortune 500 companies. The company was founded in 2009 and is headquartered in San Francisco, California. Okta is committed to providing innovative solutions that help organizations stay secure and productive in today's digital world.
Learn more about Okta
Size
5,342 employees
Market Cap
$10.5 billion
Industry
Net Income
-$266.3 million
Founded
2009
5 Year Trend
+51.9%
Revenue
$835.4 million
NASDAQ

Similar Jobs

More Jobs at Okta

More Aerospace & Defense Jobs

Find similar Staff TDI Site Reliability Engineer, Okta Federal jobs: