AstraZeneca

Generative AI Cloud Operations Engineer - Evinova

AstraZeneca$134K — $176K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Minimum 2 years of experience deploying and maintaining Generative AI agents in production.
  • Deep understanding of challenges in deploying Generative AI applications.
  • Expertise in using evaluation tools for LLMs such as Arize Phoenix or similar.
  • Strong software engineering skills in Python/TypeScript.
  • Deep knowledge of AWS services and containerization technologies like Docker and Kubernetes.

Responsibilities

  • Lead development and management of AI operations systems for clinical trials.
  • Collaborate with AI Engineers to transition projects from research to production.
  • Integrate LLM proxies and optimize RAG pipelines.
  • Develop and manage GenAI Ops systems for operational efficiency.
  • Enhance system scalability and performance through effective management.
  • Leverage modern tools and frameworks for deploying AI agents in production.

Benefits

  • 3 days a week collaborative office work in a purposefully designed space.
  • Access to a diverse and innovative team pushing boundaries in healthcare.
  • Recognition as a Top Employer for the last 10 years.
  • Support for employee flexibility in balancing work and personal commitments.
  • Commitment to diversity and inclusion in the workplace.
Full Job Description
Introduction to Role:

The Machine Learning and Artificial Intelligence Operations team (ML/AI Ops) is a newly formed platform team that will spearhead the design, creation, and operational excellence of our LLM-based agent deployments, multi-agent orchestration, and conversational AI systems pipelines to catalyze and accelerate science led innovations.

This team is responsible and accountable for the design, implementation, deployment, health and performance of all LLM-based applications. We manage ML/AI and broader cloud resources, automating operations through infrastructure-as-code and CI/CD pipelines, and ensure best-in-class operations - striving to push even beyond mere compliance with industry standards such as Good Clinical Practices (GCP) and Good Machine Learning Practice (GMLP).

As a Generative AI Cloud Operations Engineer for clinical trial design, planning, and operational optimization on our team, you will lead the development and management of AI operations systems for our trial management and optimization SaaS product. You will collaborate closely with our AI Engineers to transition projects from embryonic research into production-grade AI capabilities, utilizing advanced tools and frameworks to optimize model deployment, governance, and infrastructure performance.

This position requires a deep understanding of cloud-native agentic Generative AI deployment methodologies and technologies, AWS infrastructure, and the unique demands of regulated industries, making it a cornerstone of our success in delivering impactful solutions to the pharmaceutical industry.

Accountabilities:

Operational Excellence
  • Drive the creation of proactive capability and process enhancements that ensures enduring value creation and analytic compounding interest.
  • Design and implement resilient cloud Genereative AI agent operational capabilities to maximize our system A-bilities (Learnability, Flexibility, Extendibility, Interoperability, Scalability).
  • Drive precision and systemic cost efficiency, optimized system performance, and risk mitigation with a data-driven strategy, comprehensive analytics, and predictive capabilities at the tree-and-forest level of our Generative AI-based systems, workloads and processes.


ML/AI Cloud Operations and Engineering
  • Develop and manage GenAI Ops systems for clinical trial design, planning and operational optimization.
  • Integrate LLM proxies/routers including LiteLLM Proxy/Router or other solutions
  • Ensure proper RAG pipeline optimization and scaling
  • Integration of token usage, latency, response quality, and hallucination detection tools at a platform level.
  • Partner closely with AI Engineers and data scientists to shepherd projects from embryonic research stages into production-grade agentic Generative AI capabilities.
  • Leverage and teach modern tools, libraries, frameworks and best practices to design, validate, deploy and monitor Generative AI agents in production (including LangChain, LangGraph, Google ADK, Langfuse, DSPy, Arize Phoenix, Pinecone, Weaviate, Splunk, Grafana, Prometheus, Xray, and more)
  • Enhance system scalability, reliability, and performance through effective infrastructure and process management.
  • Ensure that any prediction we make is backed by deep exploratory data analysis and evidence, interpretable, explainable, safe, and actionable.
  • Leverage Vertex AI, Azure Foundry, OpenAI, Anthropic, and other foundation model platforms to provide reliable and stable access to LLMs


Personal Attributes:
  • Customer-obsessed and passionate about building products that solve real-world problems.
  • Highly organized and detail-oriented, with the ability to manage multiple initiatives and deadlines.
  • Collaborative and inclusive, fostering a positive team culture where creativity and innovation thrive.
  • Know when to ask for help and when to help others proactively.


Essential Skills/Experience:
  • High school diploma or GED required.
  • Minimum of 2 years of hands-on experience deploying, operating, and maintaining Generative AI agents, workflows, or applications in production environments.
  • Strong understanding of the challenges associated with production GenAI systems, including reliability, scalability, latency, cost optimization, observability, evaluation, and model performance.
  • Hands-on experience deploying agentic AI solutions using frameworks such as LangChain, LangGraph, LlamaIndex, Google ADK, Strands Agents, or similar.
  • Strong experience with LLM evaluation and observability, using platforms such as Arize Phoenix, Langfuse, Braintrust, Freeplay, or comparable tools.
  • Strong software engineering skills in Python and/or TypeScript, with experience building production-quality systems.
  • Deep expertise with AWS cloud services, including deploying and operating cloud-native AI/ML workloads.
  • Strong experience with infrastructure as code, including AWS CDK using Python and/or TypeScript.
  • Experience with containerization and orchestration technologies, including Docker and Kubernetes.
  • Strong understanding of the data science and machine learning lifecycle, with demonstrated experience moving models and AI capabilities from experimentation through production deployment and ongoing operations.
  • Experience operationalizing RAG pipelines, LLM applications, or multi-agent systems in production is strongly preferred.
  • Demonstrated ability to stay current with rapidly evolving Generative AI models, frameworks, tooling, evaluation techniques, and engineering practices.
  • Proven ability to partner effectively with AI/ML engineers, data scientists, software engineers, product teams, and other cross-functional stakeholders.
  • Strong written and verbal communication skills, with the ability to clearly document technical solutions, operational processes, and system performance.


SO, WHAT'S NEXT?

To be considered for this exciting opportunity, please complete the full application on our website at your earliest convenience - it is the only way that our Recruiter and Hiring Manager can know that you feel well qualified for this opportunity. If you know someone who would be a great fit, please share this posting with them.

Where can I find out more?
  • Explore what we're building: www.evinova.com
  • Stay connected and see our impact in action: https://www.linkedin.com/company/evinova/


Apply today to bring smarter, faster clinical trials to life!

#LI-Hybrid

Annual base salary for this position ranges from 134,855.20 to 176,997.45.
AstraZeneca is committed to providing fair and equitable compensation opportunities to all colleagues. Our compensation policies and practices have been designed to allow colleagues to progress through the salary range over time as they progress in their role. The range provided in this posting represents an offer pay range used in a majority of situations. The base pay offered will vary depending on multiple individualized factors, including the candidate's skills and experience, job-related knowledge, and other specific business and organizational needs. In some cases, offers outside the range may also be considered to address unique circumstances.

In addition, our permanent positions offer an annual Variable Pay Bonus/Short Term Incentive opportunity as well as eligibility to participate in our equity-based long-term incentive program (if applicable to role). Benefits offered for permanent roles include a competitive Flex Benefits & Retirement Savings Program, 4 weeks' paid vacation, and annual Personal Days. Fixed Term Contract/Temporary positions (excluding students) are offered a Contract Benefits Program.

About AstraZeneca

AstraZeneca is a British-Swedish multinational pharmaceutical company that specializes in the research, development, and manufacturing of prescription drugs. The company was formed in 1999 through the merger of Astra AB and Zeneca Group plc. AstraZeneca's products are used to treat a wide range of medical conditions, including cancer, cardiovascular disease, respiratory disease, and diabetes. The company has operations in over 100 countries and employs more than 76,000 people worldwide. AstraZeneca is committed to developing innovative medicines that improve the health and well-being of people around the world.
Learn more about AstraZeneca
Size
83,100 employees
Market Cap
$211.5 billion
Industry
Net Income
$3.1 billion
Founded
1999
5 Year Trend
+10.2%
Revenue
$26.6 billion
NASDAQ

Similar Jobs

More Jobs at AstraZeneca

More Information Technology Jobs

Find similar Generative AI Cloud Operations Engineer - Evinova jobs: