Steampunk

Senior LLMOps Engineer

Steampunk$145K — $185K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Ability to hold a position of public trust with the U.S. government.
  • Bachelor's, Master's, or Ph.D. in Computer Science, Machine Learning, Data Engineering, or related field.
  • 5+ years of experience in software engineering, MLOps, or cloud engineering, including 2+ years with LLM or GenAI operations.
  • Strong experience with deploying models using frameworks like Hugging Face Transformers, vLLM, or similar.
  • Proficiency in Python and operational tools such as FastAPI and PyTorch.
  • Advanced knowledge of cloud platforms like AWS, Azure, or GCP.
  • Experience with Docker, Kubernetes, and infrastructure-as-code tools.

Responsibilities

  • Architect and maintain scalable LLM and RAG pipelines.
  • Lead design and implementation of GenAI infrastructure in cloud environments.
  • Build automated evaluation systems for LLM output quality.
  • Develop CI/CD workflows for LLM applications, including dataset versioning.
  • Collaborate with AI Product Engineers to productionize LLM prototypes.
  • Integrate various components into end-to-end LLM solutions.
  • Implement observability solutions tracking performance metrics.

Benefits

  • Opportunities for mentorship and professional growth within the AI & Data Exploitation Practice.
  • Access to advanced technology and innovative AI tools.
  • Collaboration with cross-functional teams to drive generative AI projects.
  • Flexible working arrangements and support for work-life balance.
Full Job Description
Overview

We are looking for an experiencedSeniorLLMOpsEngineerto design, implement, andmaintainproduction-grade large-language-model (LLM) pipelines, deployment architectures, and monitoring systems across enterprise environments. The SeniorLLMOpsEngineer will play a critical role in operationalizing generative AI capabilities, ensuring that LLM-based applications are scalable, secure, reliable, and compliant with emerging AI risk and governance frameworks. This role spans the spectrum of model deployment, orchestration, evaluation, and optimization.

Contributions
  • Architect andmaintainscalable LLM and RAG pipelines, including model hosting, inference optimization, retrieval layers, and context management frameworks.
  • Lead the design and implementation of secure GenAI infrastructure across cloud environments, ensuring reliability, performance, and cost efficiency.
  • Build and manage automated evaluation systems that assess LLM output quality, safety, latency, and adherence to AI governance requirements.
  • Develop CI/CD workflows tailored for LLM- and GenAI-based applications, including dataset versioning, model lineage, and automated testing of prompt and model behaviors.
  • Collaborate with AI Product Engineers and Data Scientists toproductionizeLLM-based prototypes into enterprise-grade, maintainable systems.
  • Integrate vector databases, model gateways, content filters, and guardrail frameworks into end-to-end LLM solutions.
  • Implement observability and monitoring solutions that track performance metrics, hallucination rates, cost profiles, and user interaction patterns.
  • Lead troubleshooting and root-cause analysis for issues related to LLM deployment, inference performance, or pipeline reliability.
  • Stay current with emerging LLM architectures, inference optimizations, fine-tuning techniques, and relevantMLSecOpspatterns.
  • Ensure compliance with data privacy, ethical AI, and AI-governance frameworks throughout pipeline design and operations.
  • Mentor junior engineers and contribute to Steampunks AI engineering best practices, tooling, and reusable infrastructure patterns.
  • You will contribute to the growth of our AI & Data Exploitation Practice!

Qualifications
  • Ability to hold aposition of public trustwith the U.S. government.
  • Bachelors, Masters, or Ph.D. inComputer Science, Machine Learning, Data Engineering, ora related field.
  • 5+ yearsof experience in software engineering, data engineering,MLOps, or cloud engineering, with2+ yearsfocusing specifically on LLM or GenAI operations.
  • Strong experience deploying models using frameworks such asHugging Face Transformers,vLLM,TensorRT-LLM, or similar.
  • Proficiencyin Python and operational tooling such asFastAPI,PyTorch,LangChain,LlamaIndex, and vector databases(FAISS, Milvus, Pinecone, or similar).
  • Advanced knowledge ofcloud platforms(AWS, Azure, GCP) including model hosting, distributedcompute, and secure networking patterns.
  • Hands-on experience buildingCI/CD pipelines, automated testing frameworks, and environment provisioning for AI/ML workloads.
  • Experience withDocker, Kubernetes, and infrastructure-as-code(Terraform, CloudFormation).
  • Familiarity withMLSecOps, AI governance, model hardening, prompt injection defenses, and content safety monitoring.
  • Strong understanding oflogging, observability, and performance profilingfor high-throughput LLM inference systems.
  • Excellent written and verbal communication skills, with the ability to explain trade-offs and architectural decisions to technical and non-technical stakeholders.
  • Demonstrated ability to balance long-term platform thinking with hands-on operations and rapid problem solving.
  • Experience working in agile teams and using modern project management tools.

Preferred:

  • Experience building and maintaining classical ML pipelines, including feature engineering, model training, and automated retraining workflows.
  • Familiarity with ML experiment tracking and model versioning tools such as MLflow or Weights & Biases.
  • Familiarity with batch and streaming data pipeline orchestration (Airflow or similar) supporting model training workflows.
  • Experience supporting the full ML lifecycle, from data ingestion through model deployment, for classification, regression, or recommendation systems.

About steampunk

Steampunk relies on several factors to determine salary, including but not limited to geographic location, contractual requirements, education, knowledge, skills, competencies, and experience. The projected compensation range for this position is $145,000 to $185,000. The estimate displayed represents a typical annual salary range for this position. Annual salary is just one aspect of Steampunks total compensation package for employees. Learn more about additional Steampunk benefits here.

Identity Statement

As part of the application process, you are expected to be on camera during interviews and assessments. We reserve the right to take your picture to verify your identity and prevent fraud.

Similar Jobs

More Jobs at Steampunk

More Information Technology Jobs

Find similar Senior LLMOps Engineer jobs: