Staff Distributed Systems Engineer

FOMO Labs Inc

$150K — $180K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 8+ years of backend, platform, or infrastructure engineering experience.
  • Proven track record in designing and debugging distributed, high-throughput production systems.
  • Advanced skills in PostgreSQL performance optimization and management.
  • Expertise in Redis-compatible systems with a focus on sharding and replication.
  • Experience with AWS services such as ECS, RDS, and ElastiCache.
  • Proficiency in infrastructure as code, particularly with Terraform.
  • Strong programming skills in Go, TypeScript/Node.js, or similar languages.
  • Experience in designing and testing failover and disaster-recovery systems.

Responsibilities

  • Design and operate high-throughput, multi-region backend services.
  • Enhance datastore and cache performance, scalability, and reliability.
  • Implement traffic management strategies like backpressure and load shedding.
  • Reduce latency across regions and improve data accessibility.
  • Architect and test failover and disaster recovery procedures.
  • Collaborate on new feature designs for scalability from launch.
  • Educate the team on scalable system design principles.

Benefits

  • Flexible work arrangements that promote work-life balance.
  • Opportunities for professional development and continuous learning.
  • Access to advanced tools and technologies to enhance productivity.
  • Supportive team culture focused on collaboration and innovation.
Full Job Description
About the role

We are looking for a Staff Distributed Systems Engineer to own the reliability, scalability, and performance of our multi-region backend platform.

You will own critical shared infrastructure, including datastores, caches, messaging systems, and regional application services, and design systems that remain predictable during traffic surges, dependency failures, infrastructure changes, and partial regional outages.

You will also establish new failover and disaster-recovery capabilities within our stack, including defining recovery objectives and implementing the systems and testing required to recover services and data safely.

This is a hands-on engineering role with direct ownership of production systems. You will build and operate application and infrastructure components while improving data systems and observability.

Responsibilities

  • Design and operate high-throughput, multi-region services.
  • Improve datastore and cache performance, capacity, replication, and failure handling.
  • Implement backpressure, concurrency limits, load shedding, rate limiting, circuit breakers, and bounded retries.
  • Reduce cross-region latency and improve data locality.
  • Design and test service, datastore, and regional failover procedures.
  • Help architect new features to operate at scale from day one.
  • Level-up the team on how to think about scale.


Qualifications

  • 8 or more years of backend, platform, or infrastructure engineering experience, or equivalent practical experience.
  • Experience designing and debugging distributed, high-throughput production systems.
  • Strong PostgreSQL experience, including query performance, indexing, connection pooling, replication, transaction contention, and failure modes.
  • Strong experience with Redis-compatible systems such as Redis, Valkey, Dragonfly, or KeyDB, including sharding, replication, memory management, hot keys, and failure handling.
  • Experience operating services on AWS, ideally using ECS, RDS, and ElastiCache.
  • Experience with infrastructure as code, preferably Terraform.
  • Proficiency in Go, TypeScript/Node.js, or a comparable systems-oriented language.
  • Hands-on experience designing and testing failover and disaster-recovery systems, including backup restoration, replication, regional failover, and RTO/RPO validation.


Nice to have

  • Experience with NATS JetStream, Kafka, or another durable messaging system.
  • Familiarity with Datadog APM and AWS Performance Insights.
  • Experience performing live datastore or cache topology migrations.
  • Experience operating systems with bursty or unpredictable traffic.
  • Experience with financial, trading, cryptocurrency, gaming, or other high-throughput systems.

Similar Jobs

More Jobs at FOMO Labs Inc

  • Staff Distributed Systems Engineer
    $150K — $180K *
    New York, NY 10025 (New York County)
    Information Technology
    In-Person
  • Staff Backend Engineer
    $150K — $180K *
    New York, NY 10025 (New York County)
    Information Technology
    In-Person
  • Controller
    $125K — $150K *
    Remote
    Finance & Insurance
    Remote in New York City, NY
  • Controller
    $150K — $180K *
    New York, NY 10025 (New York County)
    Finance & Insurance
    In-Person
  • Brand Marketing Designer
    $80K — $95K *
    New York, NY 10025 (New York County)
    Media
    In-Person

More Information Technology Jobs

Find similar Staff Distributed Systems Engineer jobs: