Snorkel AI

Software Engineer - Platform

Snorkel AI$220K — $300K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years of experience in platform infrastructure or data systems in production environments.
  • Proficiency in Python and REST API design for internal services.
  • Background in distributed systems with hands-on experience in AWS services.
  • Familiar with data orchestration (Prefect, Airflow) and transformation (dbt) tools.
  • Understanding of data governance concepts like RBAC and data lineage.
  • Proven track record of leading engineering initiatives and influencing stakeholders.
  • Strong technical communication skills.

Responsibilities

  • Design and build agent infrastructure to accelerate workflows for teams.
  • Implement event-driven data flows ensuring reliability and recoverability.
  • Build systems to track data movement and enforce access governance.
  • Define strategy for CI/CD pipelines and transition to automated deployment.
  • Instrument services with observability tools and monitor service reliability.
  • Contribute to infrastructure cost visibility and data storage optimization.
  • Collaborate with cross-functional teams to maintain high standards in code and processes.

Benefits

  • Meaningful ownership over critical infrastructure and services.
  • Opportunity to shape foundational architecture's scalability for the future.
  • Work with a team focused on reliability and fast-paced innovation.
Full Job Description
About the Team

Snorkel's Platform organization owns the infrastructure and services that power everything at Snorkel - the pipelines, evals, access layers, event systems, governance, compute, and agent infrastructure that every product team and customer deployment depends on. We're a small team with a large surface area, and we're in the middle of a foundational architecture shift: moving from a single-database data path to a multi-source, event-driven agent first platform. The decisions being made now will define how our platform scales for years. You will be making them.

About the Role

We're looking for Platform Engineers who combine strong infrastructure chops with real backend engineering depth. You'll build and operate the systems, services, and agent infrastructure that let product teams move fast and reliably - from data access layers and event-driven pipelines to the agent first architecture that will transform our ability to scale. Your work will directly allow us to scale the amount of high quality data we are able to deliver to our customers.

We are looking to grow our team of Platform Engineers, and are hiring at multiple levels.

What You'll Do
  • Design and build agent infrastructure that allows us to safely and reliably speed up workflows from engineering and operation teams
  • Design and implement event-driven data flows using event brokers, CDC connectors, schema registries, event routing, and dead letter queues - ensuring events flow reliably and failures are visible and recoverable
  • Build the systems that track how data moves through the platform (lineage), enforce who can access what (governance and RBAC), and log what happened (auditing), including PII handling, retention policy enforcement, and audit infrastructure for enterprise and regulatory compliance
  • Set strategy and architecture for build systems, testing frameworks, and CI/CD pipelines, and drive the transition toward robust, automated continuous deployment
  • Instrument services with OpenTelemetry, define and monitor SLOs (query latency, pipeline success rates, service reliability), and build alerting that catches issues before they become incidents - you will be on-call for the systems you build
  • Contribute to infrastructure cost visibility and optimization - query cost estimation, workload right-sizing, and routing data to the most cost-effective storage tier for its access pattern
  • Collaborate with engineers, product managers, and designers to bring consistency and high standards to codebases, infrastructure, and processes

What You'll Bring
  • 5+ years building platform infrastructure, backend services, or data systems in production - you have built and operated pipelines, data access layers, distributed services, or ETL/ELT systems at scale
  • Strong proficiency in Python, and experience designing REST APIs for internal services and developers
  • Strong background in distributed systems and cloud platforms (AWS preferred) - hands-on experience with services like S3, RDS, EKS, EventBridge, and IAM, and comfort working in a Terraform-managed environment
  • Familiarity with data orchestration tools (Prefect, Airflow, or Dagster) and transformation frameworks (dbt)
  • Understanding of data governance concepts - RBAC, PII handling, audit logging, data lineage
  • Track record of leading complex engineering initiatives, influencing stakeholders, and delivering measurable impact
  • Ability to work in a fast-paced environment with strong technical communication skills
  • Fluency with modern developer tooling and a willingness to adopt new tools quickly - the team evaluates and integrates new tooling regularly to improve velocity and reliability

Nice to Have
  • Experience building shared libraries or SDKs consumed by multiple teams - versioning, backwards compatibility, migration support
  • Experience with event-driven architectures - CDC, event buses, schema registries, at-least-once delivery semantics
  • Experience with OpenTelemetry, ClickHouse, or similar observability infrastructure
  • Prior work in regulated environments (SOC 2, FedRAMP, HIPAA) where compliance requirements shaped system design
  • Experience with Ray or similar frameworks for distributed compute workloads
  • Experience in hyper-growth startup environments or scaling engineering orgs
  • Prior experience as a Tech Lead, Team Lead, or hands-on Engineering Manager

Why This Role

You'll have meaningful ownership over the infrastructure and services that every product team and customer deployment depends on. This isn't a maintenance role - you'll be making foundational architecture decisions that shape how the company scales, with a team that cares deeply about reliability, craft, and moving fast without breaking things.

Actual compensation will be determined based on factors including skills, qualifications, experience, and geographic location.

Salary range(s) for this role

$220,000-$300,000 USD

About Snorkel AI

Snorkel AI is an artificial intelligence company that provides a platform for building and managing machine learning models. The company was founded in 2019 and is headquartered in San Francisco, California. Snorkel AI's platform is designed to make it easier for developers and data scientists to create and manage machine learning models, using a technique called programmatic labeling. The company's platform is used by a number of large enterprises, including Intel, Google, and Microsoft. Snorkel AI has raised over $50 million in funding to date.
Learn more about Snorkel AI
Size
50 employees
Industry
Founded
2019

Similar Jobs

More Jobs at Snorkel AI

More Information Technology Jobs

Find similar Software Engineer - Platform jobs: