Member of Technical Staff - Lead, Storage

Modal, Inc

$160K — $200K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 7+ years of software development experience in production environments
  • 3+ years of people management experience leading engineering teams
  • Proven experience in building high-performance distributed storage or caching systems
  • Deep understanding of cloud technologies, particularly object storage and CDNs
  • Strong knowledge of operating system fundamentals including Linux and file systems
  • Experience with multi-region replication and consistency models
  • Track record in capacity planning and managing petabyte-scale data systems
  • Ability to engage in hands-on technical development alongside team management

Responsibilities

  • Lead a team to design and maintain distributed object storage for a serverless platform
  • Set technical direction for key storage primitives to support other teams
  • Own the product roadmap for major storage challenges at a large scale
  • Manage and optimize garbage collection processes for petabyte-scale datasets
  • Guide the implementation of observability and automation practices
  • Collaborate on architectural decisions impacting performance and cost
  • Engage in on-call rotations and respond to incidents to ensure system reliability

Benefits

  • Opportunity to work on cutting-edge high-performance storage systems
  • Involvement in developing solutions that handle vast amounts of data
  • Work within a collaborative environment focusing on innovation
  • Opportunity for personal growth in managing engineering teams
  • Flexible work arrangements fostering a healthy work-life balance
Full Job Description
The Role:

We are looking for a strong technical lead to guide the engineers designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. You'll lead the team responsible for the distributed object storage system that underpins every container image, volume, and checkpoint on Modal: hundreds of petabytes of data, replicated across multiple cloud object stores and a CDN, cached on local NVMe across a large fleet of workers in many datacenters, and shared peer-to-peer within each datacenter. You'll set technical direction for the primitives that other teams (filesystems, training, sandboxes) build on, balancing durability, latency, throughput, and cost. You'll own the roadmap from today's hardest problems (garbage collection at petabyte scale, active-active replication, rate limiting that protects the upstream without wasting utilization) to the architectural bets that decide what blobnet becomes: storage colocated with the GPUs, tiered writes, and capacity planning against provider limits. You'll manage a team of 3-8 engineers while staying hands-on across the stack, from local disk and page cache to distributed blob storage and garbage collection, and you'll guide the observability, automation, and on-call practices that keep the system healthy as it grows by orders of magnitude.

Requirements:
  • 7+ years of experience writing high-quality production code
  • 3+ years of direct people management experience, ideally leading a team of engineers through project planning, growth, and performance conversations
  • Experience building high-performance distributed storage or caching systems at a large scale (the more challenges you've worked through, the better)
  • Strong cloud skills, including deep familiarity with object storage (S3 or similar), CDNs, and their consistency, throughput, and cost characteristics
  • Strong knowledge of low-level operating system foundations (Linux kernel, file systems, page cache, containers, etc.)
  • Experience with replication, content addressing, and consistency models in multi-region or multi-cloud systems
  • Experience operating storage systems at scale (petabyte-scale datasets, high-throughput read/write paths, large-scale garbage collection or data migration), including owning cost and capacity planning
  • Track record of setting technical direction and driving architectural decisions across a team, and of building the primitives other teams depend on
  • Willingness to step into the thick of it with our on-call rotation and respond to production incidents
Nice-to-Haves:
  • Experience with data engineering at petabyte-scale.
  • Prior experience with Rust

Key Things the Team Is Working On:
  • P2P sharing of data across workers within a single datacenter to dramatically reduce ingress
  • Replicating data across multiple blob storage providers
  • Automating garbage collection across hundreds of petabytes of data
  • Deploying colocated storage clusters to datacenters to accelerate high-throughput customer workloads

Similar Jobs

More Jobs at Modal, Inc

  • Documentation Engineer
    $110K — $130K *
    New York, NY 10025 (New York County)
    Technical Services
    In-Person
  • Account Manager
    $110K — $130K *
    New York, NY 10025 (New York County)
    Technical Services
    In-Person
  • Account Manager
    $110K — $130K *
    San Francisco, CA 94112 (San Francisco County)
    Technical Services
    In-Person
  • Inference Engineering and Product Lead
    $175K — $210K *
    San Francisco, CA 94112 (San Francisco County)
    Enterprise Technology
    In-Person
  • Talent Programs Manager
    $100K — $120K *
    New York, NY 10025 (New York County)
    Education, Government & Non-Profit
    In-Person

More Information Technology Jobs

Find similar Member of Technical Staff - Lead, Storage jobs: