Software Engineer- BIS (Baseten Inference Stack)

Baseten

• $120K — $160K *

San Francisco, CA 94112In-Person

Information Technology

Less than 5 years of experience

Today

Be an Early Applicant

By clicking Apply, I agree with Ladders' Terms of Use and Privacy Policy

Job Overview by Ladders

Qualifications

Bachelor's, Master's, or Ph.D. in Computer Science, Engineering, or related field
Strong background in distributed systems, backend infrastructure, or platform engineering
Experience in building and operating reliable, low-latency, large-scale production systems
A focus on developer experience and usability of systems
Ability to debug complex systems across multiple technology layers
Genuine interest in inference engineering with a willingness to learn
Excellent communication and collaboration skills

Responsibilities

Develop infrastructure for large-scale distributed LLM inference
Work across the stack from user features to low-level infrastructure
Build capabilities for routing, autoscaling, scheduling, and observability
Enhance reliability, scalability, and usability of the inference stack
Collaborate with engineers to implement inference optimizations
Define best practices for testing, automation, and operations
Debug complex production systems involving Kubernetes and distributed workloads
Own projects from architecture through deployment and iteration

Benefits

100% coverage of medical, dental, and vision insurance for employees and dependents
Flexible PTO policy with a company-wide Winter Break
Paid parental leave
Fertility and family-building stipend through Carrot
Company-facilitated 401(k)
Opportunities for learning and networking within various ML startups

Full Job Description

THE ROLE

Baseten's Inference Stack team builds the distributed runtime that powers large-scale LLM inference across our platform. We operate at the intersection of distributed systems, model performance, infrastructure, and developer experience. We enable customers to deploy and operate cutting-edge LLM models with industry-leading performance, scalability, reliability, and ease of use.

As a Software Engineer on the Inference Stack team, you'll work across the stack - from the developer experience customers use to deploy models, the libraries used for features like tool calling and reasoning, all the way down to the systems we use to orchestrate deployments in Kubernetes and route traffic efficiently.

This is an ideal role for engineers who enjoy owning systems in production, solving hard integration problems, and making complex infrastructure simple and reliable for users.

EXAMPLE INITIATIVES

Blog Posts

https://www.baseten.co/blog/nvidia-dynamo-day-baseten-inference-stack/

https://www.baseten.co/blog/how-baseten-achieved-2x-faster-inference-with-nvidia-dynamo/

https://www.baseten.co/blog/how-baseten-multi-cloud-capacity-management-mcm-powers-cloud-self-hosted-and-hybr/#comparing-deployment-options-cloud-vs-self-hosted-vs-hybrid

RESPONSIBILITIES

Develop infrastructure and orchestration systems for deploying and managing large-scale distributed LLM inference
Work across the stack, from customer-facing features to low-level infrastructure components
Build platform capabilities related to routing, autoscaling, scheduling, observability, and runtime management
Improve the reliability, scalability, and usability of our inference stack
Collaborate closely with Model Performance engineers to make new inference optimizations broadly available to customers and easy to configure
Help define best practices around testing, release automation, benchmarking, and operational excellence
Debug complex production systems spanning Kubernetes, distributed runtimes, networking, and GPU workloads
Make thoughtful engineering tradeoffs balancing performance, reliability, operational simplicity, and developer experience
Own projects end-to-end: from architecture and implementation through deployment, monitoring, and iteration based on customer feedback

REQUIREMENTS

Bachelor's, Master's, or Ph.D. in Computer Science, Engineering, or a related field
Strong background in distributed systems, backend infrastructure, or platform engineering
Experience building and operating production systems where reliability, latency, and scale are first-class concerns
Strong sense of developer experience: you think about how systems are used, not just how they work
Motivated and willing to learn new languages, frameworks, and systems as needed
Ability to debug complex systems across multiple layers of the stack
Genuine interest in inference engineering. You don't need to have hands on experience but are willing to learn
Excellent communication and collaboration skills

BONUS

Experience with Kubernetes, including concepts like operators and custom resources
Prior work on Dynamo, vLLM, SGLang, TensorRT-LLM, or similar inference frameworks
Experience with distributed scheduling, autoscaling, or service orchestration
Experience operating GPU workloads in production
Familiarity with observability tooling, CI/CD systems, or release automation
Experience contributing to open-source infrastructure or ML systems

BENEFITS

Competitive compensation, including meaningful equity.
100% coverage of medical, dental, and vision insurance for employee and dependents
Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
Paid parental leave
Fertility and family-building stipend through Carrot
Company-facilitated 401(k)
Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.

Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.

* Ladders Estimates

Similar Jobs

Software Compliance Engineer
$90K — $180K *
Abbott
Milpitas, CA 95035 (Santa Clara County)
Reposted Today
Software Compliance Engineer
$90K — $180K *
Abbott
Alameda, CA 94501 (Alameda County)
Reposted Today
.Net/C# Developer, SaaS Manager
$123K — $172K *
1PASSWORD
Remote
Reposted Today
Software Engineer, GPU Performance
$147K — $211K *
Google
Sunnyvale, CA 94087 (Santa Clara County)
Today
Software Engineer - Motion & Behavioral Planning
$129K — $247K *
DiDi Labs
San Jose, CA 95123 (Santa Clara County)
Today
Software Engineer II
$90K — $130K *
Monogram Health
Remote
Reposted Today

Get Ready For Your
Next Interview

More Jobs at Baseten

Software Engineer- BIS (Baseten Inference Stack)
$120K — $160K *
San Francisco, CA 94112 (San Francisco County)
Today
Information Technology
In-Person
Executive Recruiter
$150K — $200K *
San Francisco, CA 94112 (San Francisco County)
2 weeks ago
Staffing
In-Person
Recruiting Operations Lead
$120K — $150K *
San Francisco, CA 94112 (San Francisco County)
2 weeks ago
Staffing
In-Person
Engineering Manager, Cloud Platform
$150K — $180K *
San Francisco, CA 94112 (San Francisco County)
2 weeks ago
Enterprise Technology
In-Person
Senior Manager, Cloud Platform & Site Reliability
$130K — $180K *
San Francisco, CA 94112 (San Francisco County)
2 weeks ago
Enterprise Technology
In-Person

More Information Technology Jobs

Business Development Director
$300K — $345K + $120K bonus *
Tier1 IT Services Firm
Kansas City, MO 64116 (Clay County)
6 days ago
Client Partner / Business Developemnt - Banking
$250K — $320K + $70K bonus *
IT Services Firm (client of TechLink Systems)
New York, NY 10001 (New York County)
6 days ago
Senior Data Engineer
$120K — $150K *
ECS
Remote
Today
Engineer I- Software
$70K — $95K *
Microchip Technology
Chandler, AZ 85225 (Maricopa County)
Today
Software Engineer lll - Payments Modernization
$102K — $179K *
Bank of America Corporation
Charlotte, NC 28269 (Mecklenburg County)
Reposted Today

Find similar Software Engineer- BIS (Baseten Inference Stack) jobs:

Nationwide San Francisco, CA

Software Engineer- BIS (Baseten Inference Stack)

Job Overview by Ladders

Full Job Description

Get Ready For Your Next Interview

Find similar Software Engineer- BIS (Baseten Inference Stack) jobs:

Get Ready For Your
Next Interview