Member of Technical Staff (Inference)

Artificial Analysis

$150K — $180K *
Enterprise Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 3+ years of professional experience, with at least 2 years closely working with inference providers or neoclouds
  • Strong analytical and critical thinking skills
  • Proficiency in Python and data analysis
  • Familiarity with model serving stack and inference APIs
  • Fluency in inference performance metrics and economics
  • Genuine interest and knowledge of Frontier AI

Responsibilities

  • Own performance and price coverage of serverless API inference across various providers
  • Drive new benchmarking dimensions, including cached pricing and endpoint accuracy
  • Work directly with inference providers and neoclouds for benchmarking
  • Produce industry analysis on inference performance and economics
  • Shape the roadmap of the inference benchmarking platform with engineers
  • Embrace AI-native workflows to leverage cutting-edge AI tools

Benefits

  • Influence AI development priorities with your work
  • Gain expertise in evaluating all major AI models
  • Manage relationships with leading AI labs and enterprises
  • Join a rapidly growing team at a pivotal moment
  • Receive competitive compensation including equity
Full Job Description
Job Description - Member of Technical Staff (Inference)

Location: San Francisco (on-site at our offices)

The Opportunity

Language model inference is the fastest-moving market in AI: dozens of providers, constant price and performance shifts, and billions of dollars of deployment decisions riding on independent data. Our inference benchmarks are the industry's reference point, and we're hiring a Member of Technical Staff to drive them.

You'll own coverage of the serverless inference landscape: benchmarking endpoints across quality, speed and price, extending our methodology to new dimensions like cached pricing, endpoint accuracy and agentic performance, and working directly with the inference providers and neoclouds who ship on our numbers.

What You'll Do
Benchmark the Inference Landscape: Own performance and price coverage of serverless API inference across the provider ecosystem, from frontier labs to specialist providers
Extend Our Methodology: Drive new benchmarking dimensions including cached pricing, endpoint accuracy and agentic performance, keeping our measurement ahead of how the industry deploys
Partner with Providers: Work directly with inference providers and neoclouds to benchmark their endpoints, resolve methodology questions and shape how the market measures serving performance
Analyze the Market: Produce the analysis the industry uses to understand inference performance and economics, from throughput and time-to-first-token to price-performance frontiers
Drive Product Direction: Shape the roadmap of our inference benchmarking platform together with our engineers and pillar lead
Become AI-Native: Embrace an AI-native workflow, using cutting-edge AI tools to generate leverage in a fast-changing industry and maintain our competitive edge in AI benchmarking

What We're Looking For

You should know the inference market from the inside.

Backgrounds include: engineering, product, developer relations or technical GTM roles at inference providers and neoclouds (e.g. Together AI, Fireworks, Baseten, Cerebras, Novita, Parasail, DeepInfra, Modal, CoreWeave, Lambda, Nebius, Crusoe or similar), or teams serving models in production at scale.

Required:
• 3+ years of professional experience, including at least 2 years at, or working closely with, inference providers, neoclouds or teams serving models in production
• Strong analytical and critical thinking skills
• Proficiency in Python and data analysis
• Hands-on familiarity with the model serving stack (e.g. vLLM, SGLang, TensorRT-LLM) and inference APIs across providers
• Fluency in inference performance metrics and economics: tokens per second, time to first token, throughput versus latency trade-offs, cost per token
• Genuine, demonstrable interest and knowledge of Frontier AI. We want people who have informed opinions about where AI is heading, not just people who use AI tools

Why Artificial Analysis?
Shape how AI gets built: The leading AI labs track our benchmarks and use them to guide their development priorities. Your work will directly influence the direction of AI.
Become a world expert in AI: You will evaluate every major model, across every major capability, as they are released. Very few roles offer this breadth of exposure to frontier AI.
Work with the most important players in AI: You'll manage relationships with teams at the leading AI labs and major enterprises as a trusted, independent voice.
Join at a defining moment: We're 40+ people, on track to double by end of year, backed by some of the most connected investors in AI. The people who join now will shape the product, the team, and the strategy as we scale.
Competitive compensation including equity

Similar Jobs

More Jobs at Artificial Analysis

More Enterprise Technology Jobs

Find similar Member of Technical Staff (Inference) jobs: