Rapid Ratings International

Head of Platform & Cloud Engineering

Rapid Ratings International$145K — $174K *
Enterprise Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 5-7 years of experience in cloud architecture and distributed systems with a focus on AWS and Kubernetes.
  • Proficient in implementing SOC 2 and ISO 27001 security controls in cloud environments.
  • Strong skills in systems engineering, including Linux internals and networking.
  • Expert in using infrastructure as code tools like Terraform.
  • Experience with software development in languages like Python, Go, or Bash for platform tooling.
  • Hands-on experience with AI operations, including model-hosting and orchestration infrastructure.
  • Proven track record of leading and mentoring SRE teams, especially across international borders.

Responsibilities

  • Architect and optimize high-availability, multi-region AWS infrastructure.
  • Lead the SRE function to establish robust telemetry and SLO frameworks.
  • Build and operate AI infrastructure supporting model-serving and governance frameworks.
  • Own threat detection and compliance posture for the platform.
  • Design self-service infrastructure patterns to optimize cloud cost and enhance engineering workflows.

Benefits

  • Hybrid work model with 3 days in office per week.
  • Opportunity to lead and mentor a distributed SRE team.
  • Access to cutting-edge technology and AI infrastructure projects.
  • Emphasis on operational excellence and innovation in platform engineering.
Full Job Description
Team: Platform & Cloud Engineering

Manages: SRE team

Location: Greater NYC or Boston Area - Hybrid: 3 days per week in the NYC or Quincy office

Role Overview

We are hiring a deeply technical senior leader to own the cloud platform, resilience, security posture, and AI infrastructure powering RapidRatings' products. The role blends deep systems engineering with modern SRE practice, bridging low-level cloud architecture, high-scale distributed systems, and platform engineering. As AI and developer self-service absorb routine pipeline toil, you will raise the technical bar for operational excellence, lead our SRE team, and build the paved-road infrastructure our AI systems, agents, and product squads run on.

What You'll Own
  • AWS architecture & distributed systems. Architect high-availability, multi-region AWS infrastructure optimized for scale, latency, resilience, and structural reliability.
  • Reliability & operational excellence. Lead the SRE function to establish robust telemetry, SLO frameworks, and rigorous root-cause analysis through blameless post-incident reviews. Drive self-healing automation that minimises operational overhead and mean-time-to-recovery.
  • AI infrastructure. Build and operate the model-serving layer, routing pipelines, cost attribution, and governance frameworks for our AI agents, copilots, and Platform Model Context Protocols (MCPs).
  • Security & compliance posture. Own the platform's threat detection and defense. Embed SOC 2 Type II and ISO 27001 controls directly into infrastructure code and automated guardrails.
  • Paved-road platform & FinOps. Design self-service infrastructure patterns so engineering squads ship securely and without friction, while continuously driving cloud cost optimization.


What This Role Is Not

This is not pipeline maintenance. Product developers own their own build, test, and deployment workflows, using AI tooling on top of the paved road you provide. You build the foundational platform and guardrails, not application-level scripts. We're hiring for the leverage, not the toil.

Qualifications & Technical Bar
  • Kubernetes at the core. Deep, hands-on experience running production Kubernetes (EKS) at scale - this is the backbone of our platform, not an add-on.
  • AWS mastery. Deep, hands-on experience across the native AWS stack (EKS/ECS, IAM, VPC networking, CloudFront, Lambda, RDS/DynamoDB) and infrastructure as code (Terraform / OpenTofu).
  • Systems engineering. Strong foundation in Linux internals, networking (TCP/IP, DNS, routing), performance tuning, and operating distributed systems at scale.
  • Security & compliance. Direct experience implementing and maintaining SOC 2 and ISO 27001 technical controls within cloud-native environments.
  • Software development. Proficiency in Python, Go, or Bash for building platform tooling, custom telemetry, and automated remediation.
  • AI/LLM operations. Hands-on exposure to model-hosting patterns, vector databases, API gateways, and LLM orchestration infrastructure.
  • Leadership & mentorship. Track record of managing distributed SRE/infrastructure teams, including international squads, while remaining directly involved in architectural design.


Who You Are
  • A natural problem solver who stays curious, works logically, and digs past symptoms to root cause.
  • You treat AI as a strong collaborator, building the platforms and guardrails that lift the whole engineering group rather than only your own output.
  • You set technical direction and carry others with you through standards, design review, documentation, and mentoring.
  • You communicate clearly across every tier, from technical specialists to the executive group, with meticulous attention to detail.
  • You bring a calm, positive attitude under pressure, including during production incidents and against tight deadlines.


Role Shape
  • Balance: roughly 50% technical architecture and systems design, 50% team management, SRE strategy, and mentorship. Deeply technical, still in the architecture.
  • Scope: direct manager for the SRE team; the authoritative technical leader setting cloud and platform standards across all engineering squads.


*Salary*: $145,000 - $174,500

About Rapid Ratings International

Rapid Ratings International is a financial services company that provides financial health ratings and risk management solutions to businesses and investors. The company's proprietary Financial Health Rating (FHR) system is designed to help clients assess the financial stability of their business partners and make informed decisions about credit risk. Rapid Ratings International was founded in 2007 and is headquartered in New York City.
Learn more about Rapid Ratings International
Size
100 employees
Industry
Founded
2001

Similar Jobs

More Enterprise Technology Jobs

Find similar Head of Platform & Cloud Engineering jobs: