Ampere Computing

AI Accelerator, Software Principal Engineer- Full-Stack

Ampere Computing$195K — $292K *
Consumer Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in Computer Science, Mathematics, or related field with 8 years of experience; or Master's degree with 6 years.
  • Hands-on experience with PyTorch.
  • Experience deploying AI models to edge, cloud GPU, or specialized accelerators.
  • Familiarity with graph optimization and performance tuning for models.
  • Experience with middleware and distributed communication layers.
  • Understanding of hardware/software interaction in AI systems.
  • Experience collaborating with hardware or platform partners.

Responsibilities

  • Deploy and validate AI models across local, on-prem, or edge platforms, focusing on runtime optimization.
  • Develop and tune inference graph transformations using torch.export workflows.
  • Collaborate with teams to enhance model execution efficiency on various AI accelerators.
  • Integrate model inference pipelines into scalable services with batching and streaming capabilities.
  • Build and maintain middleware and communication layers for modular integrations.
  • Support long-term platform development for production-ready inference services.

Benefits

  • Premium medical, dental, and vision insurance along with income protection and a 401K retirement plan.
  • Unlimited Flextime and 10+ paid holidays for better work-life balance.
  • Healthy snacks and energizing drinks available in the workplace.
Full Job Description
Description

About the Role:

We are looking for an engineer with strong experience in PyTorch-based AI deployment, accelerated inference execution, and systems integration across software components. The role involves working on temporal and multi-modal workloads, optimizing execution on target platforms, and building infrastructure to run AI models reliably in production environments.

What You'll Achieve:
  • Deploy and validate different AI models across supported inference environments (local, on-prem, or edge/accelerated platforms), optimizing runtime performance and reliability.
  • Develop and tune inference graph transformations, including torch.export-based graph workflows.
  • Collaborate with platform and infrastructure teams to improve model execution efficiency on supported AI accelerators and compute environments.
  • Integrate model inference pipelines into scalable services, including batching, streaming, and runtime orchestration.
  • Build and maintain middleware and communication layers to support modular, scalable system integration.
  • Support long-term platform development for end-to-end inference services and production readiness.

Relevant Technical Areas:

  • Temporal model architectures
  • Multi-frame or sequence embeddings (e.g., video/text sequences)
  • Attention-based models
  • Multi-modal workloads (e.g., text, imaging, and other feature modalities)
  • Automated labeling, evaluation, and validation workflows
  • CPU/runtime performance optimization for system components
  • Publish-subscribe middleware and distributed communication systems

About You:
  • Bachelors degree in Computer Science, Mathematics or a related technical field & 8 years of related experience; or Master's degree & 6 years
  • Strong hands-on experience with PyTorch
  • Experience deploying AI models to accelerated or constrained environments (e.g., edge, cloud GPU, or specialized accelerators)
  • Familiarity with graph optimization and model-performance tuning
  • Experience working with middleware, messaging, or distributed communication layers
  • Good understanding of hardware/software interaction in AI systems
  • Experience collaborating with hardware or platform partners

What We'll Offer:

At Ampere we believe in taking care of our employees and providing a competitive total rewards package that includes base pay, cash long-term incentive, and comprehensive benefits. The full base pay range for this role is between $195,000 and $292,000. Our benefits include health, wellness, and financial programs that support employees through every stage of life.

Benefit highlights include:
  • Premium medical insurance, dental insurance, vision insurance, as well as income protection and a 401K retirement plan, so that you can feel secure in your health and financial future.
  • Unlimited Flextime and 10+ paid holidays so that you can embrace a healthy work-life balance.
  • A variety of healthy snacks, energizing espresso, and refreshing drinks to keep you fueled and focused throughout the day.

And there is much more than compensation and benefits. At Ampere, we foster an inclusive culture that empowers our employees to do more and grow more. We are excited to share more about our career opportunities with you through the interview process. Our benefits include health, wellness, and financial programs that support employees through every stage of life.

#LI-CB1

#LI-DR

#LI-Hybrid

About Ampere Computing

Ampere Computing is a semiconductor company that designs and manufactures high-performance processors for cloud and edge computing. The company's processors are based on the Arm architecture and are optimized for power efficiency and performance. Ampere Computing was founded in 2017 by former Intel president Renee James and is headquartered in Santa Clara, California.
Learn more about Ampere Computing
Size
200 employees
Industry
Founded
2017

Similar Jobs

More Jobs at Ampere Computing

More Consumer Technology Jobs

Find similar AI Accelerator, Software Principal Engineer- Full-Stack jobs: