OpenAI

Inference Engineer, Robotics

OpenAI • $150K — $180K *
Consumer Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 5-7 years of experience in model performance optimization at the inference layer
  • Strong background in kernel-level systems and data movement
  • Proven skills in low-level performance tuning
  • Ability to work collaboratively with research and product teams
  • Experience in scaling AI systems for multimodal workloads
  • Comfortable navigating ambiguity and setting technical direction

Responsibilities

  • Improve model serving and inference performance for robotics research
  • Optimize systems for better throughput and reliability
  • Collaborate with cross-functional teams to enhance model scalability
  • Design and enhance serving infrastructure for robotics needs

Benefits

  • Hybrid work model with 3 days in the office each week
  • Relocation assistance for new hires
Full Job Description
About the Role

We're looking for a GPU Inference Engineer to contribute to improvements in model serving efficiency for our Robotics research. This is a high-impact role where you'll drive initiatives to optimize inference performance and scalability. You'll also be engaged in model design, to help assist our researchers in developing inference-friendly models.

This role is critical to scaling the team's broader goals - it will directly enable leadership to focus on higher-leverage initiatives by building a stronger technical foundation.

In this role you will:
  • Perform engineering efforts focused on improving model serving, inference performance, and system efficiency
  • Drive optimizations from a kernel and data movement perspective to improve system throughput and reliability
  • Partner closely with research and product teams to ensure our models perform effectively at scale
  • Design, build, and improve critical serving infrastructure to support Robotics growth and reliability needs

You might thrive in this role if you:
  • Have deep expertise in model performance optimization, particularly at the inference layer
  • Have a strong background in kernel-level systems, data movement, and low-level performance tuning
  • Are excited about scaling high-performing AI systems that serve real-world, multimodal workloads
  • Can navigate ambiguity, set technical direction, and drive complex initiatives to completion

This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees.

About OpenAI

OpenAI is an artificial intelligence research laboratory consisting of the for-profit corporation OpenAI LP and its parent company, the non-profit OpenAI Inc. The company was founded in 2015 by a group of technology leaders, including Elon Musk, Sam Altman, Greg Brockman, Ilya Sutskever, and John Schulman. OpenAI's mission is to develop and promote friendly AI for the betterment of humanity. The company has developed a number of cutting-edge AI technologies, including GPT-3, a language processing system that can generate human-like text. OpenAI has received funding from a number of high-profile investors, including LinkedIn co-founder Reid Hoffman and venture capitalist Peter Thiel.
Learn more about OpenAI
Size
100 employees
Industry
Founded
2015

Similar Jobs

More Jobs at OpenAI

More Consumer Technology Jobs

Find similar Inference Engineer, Robotics jobs: