Job Description:We are currently looking for a Site Reliability Engineer to join our Platform Engineering team in New York, NY.
About the RoleJoin our Platform Engineering team as a Site Reliability Engineer (SRE), where you will help operate and improve the reliability of a large-scale, hybrid infrastructure spanning on-premises colocation datacenters and multi-cloud environments (Azure, AWS, GCP). You will work alongside experienced engineers supporting a broad technology estate - including enterprise middleware, databases, storage, and growing AI workloads - applying DevSecOps and GitOps practices to keep our platforms performant, secure, and resilient. This is an excellent opportunity for a recent graduate eager to build deep, hands-on infrastructure and reliability engineering experience in a complex, real-world environment.
Responsibilities- Monitor production systems and respond to incidents using enterprise observability tooling, contributing to alerting, dashboards, and SLO tracking.
- Assist in operating and maintaining on-premises colocation and cloud infrastructure across compute, storage, networking, and enterprise middleware layers.
- Support multi-cloud operations across Azure, AWS, and GCP, including compute, networking, storage, and identity management.
- Contribute to automation and Infrastructure-as-Code initiatives to reduce toil and improve consistency across environments.
- Participate in DevSecOps and GitOps workflows, including CI/CD pipelines and policy enforcement.
- Support cloud security posture management efforts, helping remediate misconfigurations and vulnerabilities across cloud environments.
- Participate in on-call rotation and contribute to blameless post-incident reviews.
- Document runbooks, procedures, and system configurations to support team knowledge sharing.
Required Qualifications, Capabilities, and Skills- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field - or equivalent hands-on experience.
- Foundational understanding of Linux and/or Windows server operating systems.
- Basic knowledge of networking concepts: TCP/IP, DNS, DHCP, VLANs, and firewalls.
- Exposure to at least one scripting or programming language (Python, Bash, or PowerShell).
- Familiarity with cloud computing concepts across one or more of Azure, AWS, or GCP.
- Strong analytical and problem-solving skills with a desire to learn in a fast-paced environment.
- Effective written and verbal communication skills with the ability to document technical processes clearly.
Preferred Qualifications, Capabilities, and Skills- Experience in infrastructure, cloud, DevOps, or SRE-related roles.
- Hands-on exposure to observability or monitoring tools (Datadog, Prometheus, Grafana, or similar).
- Familiarity with IaC tools such as Terraform or Ansible and GitOps workflows.
- Experience with containerization or orchestration (Docker, Kubernetes).
- Cloud certifications (AWS Cloud Practitioner, Azure Fundamentals AZ-900, GCP Associate, or equivalent).
- Familiarity with cloud security posture or compliance concepts.
The anticipated starting salary range for individuals expressing interest in this position is $120,000 - $150,000 per year. Placement within this range is dependent upon level of experience, location and other factors. This position is eligible for annual incentive compensation which will be a part of the total compensation. Total compensation for this position will be competitive with the market.
#LI-AH1