FactSet

Principal Site Reliability Engineer (Kubernetes Required) - Hybrid

FactSet • $190K — $220K *
Enterprise Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 8+ years of experience in systems reliability, scalability, and performance
  • Hands-on experience with Kubernetes deployment and management
  • Strong knowledge of core Kubernetes concepts
  • Proficiency in cloud platforms such as AWS, GCP, or Azure
  • Experience with CI/CD tools like GitHub Actions or ArgoCD

Responsibilities

  • Monitor and enhance the reliability and availability of production systems
  • Resolve incidents and conduct post-mortems for system failures
  • Define and track Service Level Objectives (SLOs) and Indicators (SLIs)
  • Collaborate with dev teams to integrate reliability into system designs
  • Design automations to optimize operational efficiency
  • Participate in an on-call rotation for critical systems support
  • Contribute to capacity planning and performance optimization efforts

Benefits

  • Opportunities for professional development and continuous learning
  • Collaborative and supportive work culture
  • Flexible working arrangements
  • Access to cutting-edge technology and tools
  • Commitment to employee well-being and mental health
Full Job Description
About the Role

We are looking for a skilled and motivated Principal Site Reliability Engineer to join our team. In this role, you will be responsible for ensuring the reliability, scalability, and performance of our systems and services. You will work closely with development and operations teams to build and maintain robust infrastructure, automate processes, and drive engineering best practices. 

 

Key Responsibilities 

  • Monitor, maintain, and improve the reliability and availability of production systems 
  • Respond to and resolve incidents, conducting thorough post-mortems to prevent recurrence 
  • Define and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs) 
  • Collaborate with development teams to build reliability into services from the ground up 
  • Design and implement automation to reduce toil and improve operational efficiency 
  • Participate in an on-call rotation to support critical systems 
  • Contribute to capacity planning and performance optimization efforts 
  • Document systems, processes, and runbooks to support the wider team 

 

 

Minimum Experience:

  • 8+ years’ experience ensuring the reliability, scalability, and performance of our systems and services

 

 

Required Technical Skills 

Kubernetes (Required) 

  • Hands-on experience deploying, managing, and troubleshooting workloads in Kubernetes 
  • Strong understanding of core Kubernetes concepts including Pods, Deployments, Services, ConfigMaps, and Ingress 
  • Experience with Kubernetes cluster management and administration 
  • Familiarity with Helm for application packaging and deployment 
  • Understanding of Kubernetes networking, storage, and security best practices 

 

 

Additional Technical Skills 

  • Cloud Platforms: (e.g. AWS, GCP, Azure) 
  • CI/CD Tooling: (e.g. GitHub Actions, ArgoCD, Harness) 
  • Monitoring & Observability: (e.g. Prometheus, Grafana, Coralogix, OpenTelemetry) 
  • Infrastructure as Code: (e.g. Terraform, Pulumi) 
  • Config Management: (e.g. Ansible, Puppet, Chef) 
  • Programming/Scripting: (e.g. Python, Go, Bash) 

 

 

Soft Skills & General Requirements 

  • Strong problem-solving and analytical skills with a methodical approach to troubleshooting 
  • Excellent communication skills with the ability to collaborate across technical and non-technical teams 
  • A proactive mindset with a focus on automation and continuous improvement 
  • Ability to work effectively under pressure, particularly during incident response 
  • Commitment to a blameless culture and continuous learning 

 

 

Nice to Have 

  • Experience contributing to open-source projects 
  • Familiarity with SRE principles as defined by the Google SRE handbook 
  • Previous experience in a DevOps or Platform Engineering role 

 


Education:

  • Bachelor’s degree in computer science or relevant degree.

 

The budgeted salary range for this position in the states of Connecticut and New York is $190,000.00 - 220,000.00



About FactSet

FactSet Research Systems Inc. provides integrated financial information and analytical applications to the investment community in the United States, Europe, and the Asia Pacific. The company delivers insight and information to investment professionals through its analytics, services, contents, and technologies. Its applications suite offers tools and resources, including company and industry analyses, full screening tools, portfolio analysis, risk profiles, alpha-testing, portfolio optimization, and research management solutions. The company was founded in 1978 and is headquartered in Norwalk, Connecticut.
Learn more about FactSet
Size
10,691 employees
Market Cap
$15 billion
Industry
Net Income
$380.1 million
Founded
1978
5 Year Trend
+8.6%
Revenue
$1.5 billion
NASDAQ

Similar Jobs

More Jobs at FactSet

More Enterprise Technology Jobs

Find similar Principal Site Reliability Engineer (Kubernetes Required) - Hybrid jobs: