Software Engineer (Backend-Focused)

AZX

$120K — $145K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 4+ years of backend engineering experience in distributed systems and API design.
  • Proficiency in Go, Rust, or async Python with production experience.
  • Familiarity with LLM-specific backend concerns like rate limiting and caching is a plus.
  • Exposure to Kubernetes and containerization technologies.
  • Willingness to broaden expertise across various platform areas.

Responsibilities

  • Build and maintain backend services for the LLM gateway including routing and observability.
  • Collaborate on sandboxing and isolation infrastructure for executing agent-generated code safely.
  • Support Kubernetes-based platform services' autoscaling and operational aspects.
  • Write efficient backend code in Go, Rust, or async Python using technologies like Envoy and gRPC.
  • Ensure observability of services using OpenTelemetry as the platform scales up.
  • Work collaboratively across different teams as project priorities shift.

Benefits

  • Health insurance with substantial coverage for dependents.
  • Flexible paid time off.
  • Equity options available.
  • Fully remote work culture with a team based in Seattle.
  • Opportunities for professional training and growth.
  • Part of a profitable, mission-driven company addressing AI transformation in critical sectors.
Full Job Description
Software Engineer (Backend-Focused)

About the Role

We're looking for a Staff or Senior ML Engineer to own the technical backbone of how AZX serves and evaluates models at scale. This is a high-leverage IC role spanning our inference platform - GPU scheduling, autoscaling, and serving infrastructure for vLLM/SGLang across cloud and customer-managed clusters - and the evaluation systems that tell us whether model, prompt, and agent changes that make things better.

You'll create technical direction for how AZX serves models reliably. This role suits someone who wants architectural ownership over hard ML infrastructure problems, paired with the judgment to build the guardrails that let the rest of the team move fast safely.

What you will do

You will work on software projects in client engagements, and over time, internal platform capabilities.

You will:
  • Build and maintain backend services for our LLM gateway - routing, rate limiting, key management, and observability in front of the inference fleet.
  • Contribute to sandboxing and isolation infrastructure that keeps agent-generated code safe to execute, working alongside our security-focused engineers.
  • Support Kubernetes-based platform services, including operators and autoscaling logic adjacent to our inference platform.
  • Write high-performance backend code in Go, Rust, or async Python (FastAPI/Starlette), working with infrastructure like Envoy and gRPC.
  • Instrument services with OpenTelemetry so behavior, latency, and cost stay observable as the platform scales.
  • Collaborate across the gateway, sandbox, and inference platform teams, flexing across areas as priorities shift.


Core Qualifications - Technical and foundational
  • 4+ years of experience in backend engineering fundamentals: distributed systems, API design, and production experience in Go, Rust, or async Python.
  • Familiarity with LLM-specific backend concerns (rate limiting, caching, token accounting) is a plus, though not required on day one.
  • Exposure to Kubernetes and containerization; interest in sandboxing or security is a plus.
  • Comfort working across a range of platform concerns rather than one narrow specialty - this role is intentionally broader than our specialist infra profiles.
  • Eagerness to grow into deeper specialization in gateway, sandbox, or inference infrastructure over time.

Values and Culture Qualifications
  • High emotional intelligence and a learning mindset
  • Strong collaboration skills
  • Enjoy others' success and a fun, positive environment.
  • Comfortable making decisions in the face of ambiguity and course correcting as needed.


Bonus Qualifications (not required but a huge plus)
  • Experience in both startup and enterprise environments
  • Past work in energy, real estate, utilities, climate or related fields
  • Bonus if you have experience and passion in one or more of
    • Additional web frameworks (e.g. Svelte, Vue, Angular)
    • Lower-level languages e.g. C++, Rust
    • Networking paradigms e.g. GraphQL, Websockets
    • ML capabilities e.g. Sk-learn, xgboost, Pytorch/Tensorflow/JAX, Onnx...
    • Additional database types such as graph or vector databases
    • DevOps e.g. CI/CD pipelines, Docker, Kubernetes, Terraform, Pulumi and/or Bicep
    • Generative AI e.g. prompt engineering, RAG, fine-tuning, tooling ecosystem


Compensation & benefits
  • Competitive early-stage startup compensation (based on capabilities, experience, and location)
  • Bonus eligibility
  • Health insurance with meaningful coverage for dependents
  • Flexible paid time off
  • Equity
  • Fully remote culture with a cluster of teammates in Seattle
  • Training and learning opportunities
  • Be part of a fast-growing, profitable, mission-driven company with industry leading clients tackling the massive opportunity of AI transformation in critical industries


Logistics
  • Remote but only USA/Canada
  • Must be willing to travel to Seattle area for final interview and travel 2x/year for company summits


Next steps

If this job sounds great, we'd love to hear from you. If you feel aligned to the company but don't check all these boxes, we'd still love to hear from you!

Similar Jobs

More Jobs at AZX

More Information Technology Jobs

Find similar Software Engineer (Backend-Focused) jobs: