Senior/Staff Software Engineer - Site Reliability & Infrastructure

General Intuition

• $150K — $180K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Strong fluency in Terraform for infrastructure-as-code
  • Hands-on experience with Elasticsearch for user-facing features
  • Proficient in GCP services including Kubernetes and VPC
  • Deep experience in scaling relational databases like MySQL and Postgres
  • Calm and effective incident response experience through P0 issues
  • Familiarity with CI/CD pipelines using GitHub Actions
  • Strong communication skills for issue escalation and postmortem documentation
  • Startup experience in high-growth environments

Responsibilities

  • Own the on-call rotation and manage incident responses
  • Drive postmortems to analyze and prevent recurring issues
  • Collaborate with engineering teams to address infrastructure needs
  • Scale and maintain databases efficiently under pressure
  • Implement infrastructure-as-code practices and CI/CD processes
  • Ensure the reliability of the infrastructure as the company grows
  • Monitor and troubleshoot system health and performance

Benefits

  • Impactful work on infrastructure handling billions of clips
  • Experience in a fast-paced startup culture
  • Opportunity to engage with cutting-edge technology stack
  • Contribute to meaningful projects with direct engineering impact
  • Collaborative environment with engineering teams on scaling efforts
  • Exposure to critical incident management at scale
  • Professional growth opportunities in a rapidly evolving field
Full Job Description
The Role

Medal's infrastructure handles billions of clips, video ingestion pipelines, and social features at a massive scale most engineers never get to touch. The work centers on reliability, incident response, scaling, and making sure our infrastructure keeps up with our growth. You'll own the on-call rotation, drive postmortems, and work directly with engineering teams to meet their infra needs. The right person probably came through startups and scale-ups and has been in the room when things broke at 2am, has scaled databases under pressure.
What We're Looking For
  • Infrastructure-as-code: Strong fluency in Terraform, with real experience owning infrastructure-as-code at scale
  • Elastic search depth: Hands-on experience running ES for user-facing features, not just as a log sink
  • GCP depth: Kubernetes, VPC, IAM, Cloud Logging, and the managed services ecosystem
  • Database scaling: Deep, hands-on experience scaling and sharding relational databases (MySQL, Postgres) in production
  • Incident response instincts: You can work a P0 calmly, communicate clearly under pressure, and run a postmortem that prevents recurrence
  • CI/CD: You've worked with GitHub Actions in a production environment
  • Communication (crucial!): You flag issues clearly and rapidly during incidents and lead/write actionable postmortems
  • Experience at startups: You are comfortable in an environment of rapid growth where scaling up is a priority
  • Great judgment: You know the difference between a durable, sustainable fix and a patch that buys you a week
Our Stack

Electron, React, Redux, Styled Components & other modern web-based technologies
C# and C++ for native Windows recording & more
Swift for iOS, Kotlin for Android
Java, Redis, RabbitMQ, Kubernetes for backend
Terraform, Salt, GitHub Actions, CircleCI for IaC and CI/CD

Similar Jobs

More Jobs at General Intuition

More Information Technology Jobs

Find similar Senior/Staff Software Engineer - Site Reliability & Infrastructure jobs: